Search papers, labs, and topics across Lattice.
4
0
4
1
LigBench transforms the landscape of LLM-driven research idea generation by providing a unified benchmark that aligns closely with expert evaluations.
OrthoPilot outperformed seasoned orthopaedic experts in diagnostic reasoning, achieving a 10.6% increase in management success for complex musculoskeletal cases.
Claim drift in automated research can lead to significant discrepancies, but Xcientist ensures that every generated mechanism remains accountable and traceable back to its evidential roots.
LLMs still struggle to answer questions about AI research papers, as evidenced by a new challenging dataset, AirQA, which also comes with an automated method for synthesizing training data to improve performance.