Search papers, labs, and topics across Lattice.
Harbin Institute of Technology
2
0
5
DreOPD achieves superior performance by transforming reward extrapolation into stable velocity regression, outperforming traditional methods and specialized teachers alike.
MLLMs can significantly reduce hallucinations and improve reasoning by shifting focus from linguistic shortcuts to causal visual grounding through VIGIL.