Search papers, labs, and topics across Lattice.
4
0
4
10
Align-RAG reveals that frozen TSFMs can dynamically leverage retrievals without any learned parameters, outperforming traditional methods by a significant margin.
Pixel-Space Diffusion Transformers could redefine high-fidelity image generation by integrating visual understanding and generation in a unified model.
None of the 18 multimodal large language models audited are order-invariant, with flip rates revealing a staggering sensitivity to input ordering that challenges current evaluation practices.
AI-generated videos may look realistic, but HumanScore reveals they often fail at biomechanical fidelity, suffering from jitter, anatomically implausible poses, and motion drift.