Search papers, labs, and topics across Lattice.
Affiliation:
2
0
3
16
OmniUE achieves an unprecedented 83.7% improvement in visual-interactive benchmarks, setting a new standard for multimodal embedding performance.
MMHNet proves you can train a video-to-audio model on short clips and have it generalize to generate coherent audio for videos over 5 minutes long.