Search papers, labs, and topics across Lattice.
2
0
3
Achieving up to 48.3% better image quality and 8.9% higher accuracy in face recognition, MTVDiff revolutionizes thermal-to-visible face translation through advanced multimodal integration.
Generating realistic human-object interaction videos from text, images, audio, *and* pose is now possible, opening the door to automated content creation workflows.