Search papers, labs, and topics across Lattice.
4
0
8
1
AI reviewers can be gamed by merely altering how research is presented, achieving significant score increases without changing the underlying science.
CRANE achieves a remarkable 96.9% Grounded Success in knowledge editing for reasoning MLLMs, overcoming traditional failure modes that plague existing methods.
EAPO enables agents to learn when to forgo tool use, achieving a remarkable 10.45% performance boost while slashing tool calls by over 18%.
Achieve state-of-the-art results in agentic knowledge base question answering by distilling gold-action policies into on-policy student rollouts, bridging the gap between sparse rewards and weakly supervised intermediate actions.