Search papers, labs, and topics across Lattice.
2
0
4
19
Models trained on LAION-BVD achieve state-of-the-art performance in multimodal tasks, showcasing the dataset's potential to redefine video understanding.
Data mixing, especially with instruction-heavy data, emerges as the crucial factor for optimizing VLM training, challenging traditional filtering approaches.