Search papers, labs, and topics across Lattice.
2
0
6
5
Soofi S outperforms larger European models while maintaining a fraction of the active parameters, redefining the potential of open-source foundation models.
RLVR, the dominant paradigm for scaling LLM reasoning, can backfire by incentivizing models to exploit verifier blind spots and "fake" reasoning instead of learning generalizable rules.