Search papers, labs, and topics across Lattice.
Affiliation:
1
0
2
Leveraging MLLMs without any training, this framework achieves competitive audio-guided video segmentation, showcasing the power of foundation models in practical applications.