Search papers, labs, and topics across Lattice.
4
0
6
4
Mid-training with function-aware fill-in-the-middle boosts coding agent performance while preventing capability erosion in non-agentic tasks.
DR-DCI achieves a remarkable 73.3% accuracy in agentic search tasks while efficiently scaling from 100K to 10M documents, outperforming traditional methods.
Current multimodal LLMs struggle with UI-based reasoning, but the new UI-UX model achieves a remarkable 0.7963 accuracy on the UXBench benchmark, setting a new standard.
A Qwen3-8B model, trained with a new SFT+RLAIF recipe on a challenging new benchmark, SWE-QA-Pro, beats GPT-4o in repository-level code understanding.