Search papers, labs, and topics across Lattice.
University of Trento
3
0
5
1
Every team in the generation task produced at least one report deemed the best by human annotators, highlighting a leap in multimodal generation quality.
MLLMs often struggle with reasoning due to a failure in dynamic cross-modal coordination, but DyCo-RL fixes this by optimizing attention shifts for better performance.
End-to-end autonomous driving can ditch expert demonstrations and still achieve state-of-the-art performance, thanks to a risk-aware world model that learns to predict and avoid hazardous outcomes.