Search papers, labs, and topics across Lattice.
5
0
5
4
Instruction-following in robot manipulation can be rigorously tested with InstructMove, revealing the true capabilities of VLA models beyond visual cues.
Stochastic modality masking during training can boost bimanual robotic manipulation success rates by over 30% without complex architectural changes.
Achieving a staggering 96.5% human acceptance rate, EmbodiedGen V2 transforms how we create and utilize 3D environments for embodied AI training.
Achieving high-fidelity 3D scene reconstruction from monocular video, ManiSplat enables robots to interact with their environments in a more controllable and realistic manner.
Forget redrawing diagrams by hand: VFIG, a new vision-language model, can automatically convert rasterized figures into editable SVGs with near GPT-5.2 quality.