Search papers, labs, and topics across Lattice.
RynnBrain 1.1 introduces a family of embodied foundation models with scales of 2B, 9B, and 122B-A10B, enhancing capabilities in embodied perception, spatial reasoning, localization, and planning through a unified spatio-temporal framework. The model family incorporates contact-point prediction and native 3D grounding, leading to improved alignment with robotic manipulation tasks. Notably, the 122B-A10B model surpasses all evaluated competitors on multiple benchmarks, and real-robot experiments demonstrate superior performance over existing models in multi-task scenarios.
RynnBrain 1.1 not only outperforms all competitors in embodied cognition tasks but also redefines how robots can be trained for complex manipulation through innovative 3D grounding techniques.
We present RynnBrain 1.1, a family of embodied foundation models spanning 2B, 9B, and 122B-A10B scales. Trained with a unified spatio-temporal and physically grounded framework, RynnBrain 1.1 supports embodied perception, spatial reasoning, localization, and planning. Compared with RynnBrain 1.0, it further introduces contact-point prediction across the model family and native 3D grounding for the 2B and 9B models, yielding representations and outputs that are more directly aligned with robot manipulation. We also develop RynnBrain-VLA with a unified cross-embodiment action space and embodiment-specific masking, and deploy it on Unitree G1, Astribot-S1, and Tianji-Wuji. RynnBrain 1.1 achieves strong results on embodied cognition, localization, and 3D grounding, with the 122B-A10B model outperforming all evaluated proprietary and open-source models on VSI-Bench, MMSI, and RefSpatial-Bench. Real-robot experiments show that RynnBrain-initialized policies outperform Qwen-based and representative generalist VLAs, while joint multi-task and multi-embodiment training improves process scores and success rates over per-task training.