Search papers, labs, and topics across Lattice.
2
0
3
0
Spurious reasoning persists even after an end-of-think token is injected, complicating the transition to answering in large reasoning models.
Informative tokens can be selectively unlearned without sacrificing model performance, thanks to a novel entropy-based weighting mechanism.