Search papers, labs, and topics across Lattice.
2
0
5
0
By jointly training a keyframe sampler with an MLLM, MSJoE achieves state-of-the-art accuracy in long-form video understanding while significantly reducing computational cost.
Stop paying for verbose overthinking: BFS-PO slashes LRM output length while simultaneously boosting accuracy.