Search papers, labs, and topics across Lattice.
2
0
3
7
Real-world coding tasks can be evaluated without risk of contamination, thanks to a unique reverse-engineering approach that ensures task prompts are untraceable.
A Qwen3-8B model, trained with a new SFT+RLAIF recipe on a challenging new benchmark, SWE-QA-Pro, beats GPT-4o in repository-level code understanding.