Search papers, labs, and topics across Lattice.
4
0
7
Short-context models can achieve superior reasoning performance by leveraging long-context teacher models through innovative token alignment and training strategies.
WIFA reduces harmful refusal while minimizing benign over-refusal, achieving a remarkable drop in over-refusal rates from 25.7% to 17.4%.
Even the best LLMs struggle with Olympiad-level combinatorics, achieving only 65.4% on a benchmark designed to expose their reasoning limitations.
Draft-OPD accelerates inference by over 5x while improving speculative decoding accuracy, transforming how draft models learn from target feedback.