Search papers, labs, and topics across Lattice.
4
0
5
1
VGIF-Score reveals that current video generation models struggle with complex instructions, providing a diagnostic lens to pinpoint where they succeed or fail.
MMBench-Live achieves a high answer correctness rate while updating benchmarks at a fraction of the cost and time, revolutionizing how we assess VLMs.
Current MLLMs can't find the lies hidden in their long image captions, struggling to pinpoint specific hallucinated words within detailed narratives.
You can now get state-of-the-art hepatocellular carcinoma diagnosis and captioning from whole slide images using a new MLLM with a topology-aware attention mechanism.