Search papers, labs, and topics across Lattice.
3
0
5
Hy-Embodied-VLM-1.0 outperforms its predecessor by 8.4% while activating only a fraction of the parameters, redefining efficiency in embodied agents.
Speech QA performance peaks at 4.17 Hz, challenging the assumption that higher frame rates always yield better reasoning outcomes.
Achieving high-fidelity audio generation with just four sampling steps, AudioX-Turbo dramatically cuts inference costs while enhancing performance across multimodal tasks.