Search papers, labs, and topics across Lattice.
1
0
3
VLMs struggle with robot task evaluations, achieving only 0.77 mean balanced accuracy, and even fine-tuned models often underperform compared to general-purpose counterparts.