Search papers, labs, and topics across Lattice.
The paper introduces Athena-Brain-8B, an 8 billion parameter language model designed to function as an efficient on-device brain for embodied intelligence. It addresses the challenge of balancing general-purpose intelligence with specialized embodied capabilities through a multi-stage post-training pipeline that includes General Supervised Fine-Tuning, General Reinforcement Learning, Embodied Expert training, and Model Merge. Experimental results reveal that Athena-Brain-8B not only matches the performance of larger models on general benchmarks but also excels in embodied evaluations while generating more concise responses.
Compact models like Athena-Brain-8B can outperform larger counterparts in embodied tasks while maintaining strong general intelligence.
Large language models (LLMs) have demonstrated remarkable capabilities in language understanding, reasoning, and world knowledge. As embodied agents become increasingly capable, there is a growing demand for compact models that can serve as an on-device brain, preserving the broad general intelligence of LLMs while enabling effective high-level interaction with embodied environments. Existing approaches, however, often prioritize either general-purpose intelligence or specialized embodied capabilities, making it challenging to satisfy both requirements within a single model. We present \textbf{Athena-Brain-8B}, an 8B LLM designed to serve as an on-device brain for embodied intelligence for embodied intelligence. Through a multi-stage post-training pipeline consisting of General Supervised Fine-Tuning, General Reinforcement Learning, Embodied Expert training, and Model Merge, Athena-Brain-8B maintains strong general capabilities while acquiring strong high-level embodied interaction capabilities and generating concise responses for efficient embodied interaction. Experimental results demonstrate the effectiveness of Athena across both general and embodied evaluations. Compared with the corresponding Qwen3-8B thinking model, Athena-Brain-8B achieves comparable performance on general language and reasoning benchmarks while generating substantially shorter responses. On in-domain embodied benchmarks, Athena-Brain-8B consistently outperforms models of similar scale and surpasses several substantially larger frontier models evaluated zero-shot, demonstrating that compact language models can effectively integrate strong general intelligence with embodied capabilities.