Search papers, labs, and topics across Lattice.
3
0
5
11
Game solvers can teach LLMs how to make better decisions in long-horizon tasks by providing actionable turn-level feedback, leading to superior performance in complex environments.
Clinician-centered evaluations reveal that later active learning models are preferred, underscoring the importance of usability in ultrasound AI systems.
Agent-as-a-Judge can outperform LLM-as-a-Judge in complex environments, but still struggles to reliably verify agent behavior, revealing a critical gap in current LLM-based agent evaluation.