Search papers, labs, and topics across Lattice.
2
0
3
0
T2I models struggle significantly with spatial instructions based on object orientation, achieving only 44.3% accuracy on frame-of-reference prompts.
VLMs can now reason about temporal inconsistencies in video deepfakes, thanks to a new benchmark that moves beyond static artifact detection.