Search papers, labs, and topics across Lattice.
Affiliation:
1
0
2
Discarding "easy" examples where the base model already follows glossary constraints yields an 11-point accuracy leap at fixed data volume, outperforming RL-based alignment methods with standard supervised fine-tuning alone.