Search papers, labs, and topics across Lattice.
Zhejiang University
4
0
7
13
Scaling zero RL to a trillion parameters reveals that models can spontaneously develop advanced cognitive behaviors, making traditional heuristics obsolete.
Harness VLA boosts the performance of frozen VLA models by 38.6 percentage points on challenging manipulation tasks without the need for finetuning.
Even frontier models like Claude Sonnet 4.6 stumble when asked to infer user preferences and proactively assist in mobile tasks, achieving less than 50% success despite excelling at explicit task execution.
LLMs can escape the trap of confidently wrong reasoning by co-evolving a generator and verifier from a single model, bootstrapping each other to break free from flawed consensus.