Search papers, labs, and topics across Lattice.
4
0
7
0
Executable vector graphics enable MLLMs to achieve human-like spatial reasoning through a structured visual workspace.
Models may score well on benchmarks but often fail to meet strict perceptual requirements, revealing a hidden brittleness in multimodal evaluations.
Forget noisy pseudo-labels: SpatialEvo unlocks self-supervised 3D spatial reasoning by generating perfectly accurate training data directly from scene geometry.
Forget fine-tuning: DM0 shows that pretraining a VLA model from scratch on diverse embodied and non-embodied data leads to SOTA performance in physical AI tasks.