Search papers, labs, and topics across Lattice.
2
0
4
0
This work shows that specialist optimization implicitly selects from this latent trajectory space, and establishes a new view of specialist training: when gold reasoning is absent, tuning choices directly control the latent supervision passed to downstream models.
The results suggest that fine-tuning is not just about how much a model changes, but how that change is spent, and that changing the accessible directions can qualitatively alter the outcome of fine-tuning.