Search papers, labs, and topics across Lattice.
3
0
6
8
FaithEyes reveals that self-judging mechanisms in VLMs can drastically improve tool use fidelity, leading to more reliable multimodal reasoning.
MemTrain reveals that self-supervised memory training can outperform traditional reinforcement learning approaches in enhancing LLMs' reasoning capabilities.
TrOPD stabilizes on-policy distillation by ensuring reliable teacher supervision, leading to consistent performance improvements over existing methods.