Search papers, labs, and topics across Lattice.
Shanghai AI Laboratory
3
0
5
5
Despite LLMs excelling at identifying reviewer concerns, they falter in verifying if revisions truly resolve those issues, with the best achieving only a 0.501 score in evidence-based checks.
Achieving trillion-parameter performance with just 35 billion parameters by scaling agent horizons reveals a new frontier in model efficiency.
Agents-K1 transforms how we extract and reason about scientific knowledge, achieving superior performance in multi-hop reasoning tasks compared to existing methods.