Search papers, labs, and topics across Lattice.
4
0
8
13
FlashMorph reveals that optimizing layer selection in hybrid attention models can drastically improve efficiency while maintaining performance, outperforming existing heuristic methods.
Seedance 2.0 leapfrogs existing models by unifying multi-modal inputs (text, image, audio, video) into a single architecture for generating high-quality, longer-duration audio-video content.
LRMs already know when to stop reasoning, but current sampling methods are holding them back.
Stop overfitting your reward model: R2M leverages real-time policy feedback to dynamically align the reward model with the evolving policy distribution, reducing reward overoptimization in RLHF.