Search papers, labs, and topics across Lattice.
Nanjing University of Science and Technology
4
0
4
5
Self-evolving coding agents can revolutionize software development by learning from past interactions, but they also face significant challenges in reliability and safety.
MultiFixer repairs 420 bugs, including complex multi-hunk cases, establishing a new benchmark in Automated Program Repair.
LLMs can be taught to avoid repeating past mistakes in vulnerability repair, boosting performance by up to 39% over state-of-the-art methods.
Forget finetuning: carefully chosen context can boost LLM performance on software engineering tasks by up to 33%, and CL4SE provides the benchmark to prove it.