Search papers, labs, and topics across Lattice.
3
0
3
SER's innovative approach to grounding video reasoning demonstrates a 3.0-point leap in accuracy by integrating semantic verification into the reward structure.
Efficient context handling in video tasks can elevate multimodal models to new heights of agency and reasoning capability.
Teacher privilege in multimodal reasoning is redefined, showing that visually grounded cues can lead to superior performance in on-policy distillation.