Search papers, labs, and topics across Lattice.
4
0
7
27
Landmark bias can lead to significant inaccuracies in geo-localization, but HoloGeo effectively mitigates this issue through evidence-driven reasoning, outperforming existing models.
Forget unimodal tasks鈥擴niM throws down the gauntlet for truly unified multimodal AI, demanding models juggle any combination of text, image, audio, video, code, documents, and 3D inputs and outputs in a single, interleaved stream.
Overcome the scarcity of 4D training data by cleverly borrowing spatial understanding from 3D models and temporal dynamics from video models.
Achieve SOTA joint audio-video generation with JavisDiT++ using just 1M public training examples, rivaling performance of models trained on proprietary datasets.