Search papers, labs, and topics across Lattice.
Independent Researcher
2
0
3
Self-conditioning on verified trajectories boosts reinforcement learning performance by over 8%, revealing the power of internal feedback in credit assignment.
LLM agents can get 18% better at tasks by co-evolving their skills and tools, instead of learning them separately.