Search papers, labs, and topics across Lattice.
1
0
3
0
Image-generation models can outperform text-output VLMs in spatial tasks when answers are expressed directly in pixel space, revealing a critical interface mismatch in current evaluation methods.