Search papers, labs, and topics across Lattice.
2
0
6
Achieving a 90% boost in prefill throughput for MoE models could redefine the efficiency of large-scale language model serving.
VLMs can be significantly boosted on embodied tasks by mid-training on a carefully curated subset of VLM data that is highly aligned with the VLA domain, rivaling the performance of much larger models.