Search papers, labs, and topics across Lattice.
4
1
5
22
Critical tokens can make or break code generation, and ACToR ensures they receive the precise context needed to succeed.
PTA-IRT reveals that incorporating detailed problem-solving trajectories can dramatically enhance the evaluation accuracy of software engineering agents.
Even the top-performing LLM struggles with cross-file reasoning, achieving only 69.1% accuracy on a new benchmark designed to reflect real-world software development challenges.
PhoenixRepair redefines how software agents explore repair strategies, achieving a 76% resolution rate by leveraging multi-location sampling and iterative refinement.