Search papers, labs, and topics across Lattice.
Zhejiang University
7
0
7
HierCAD achieves unprecedented fidelity in CAD generation by aligning structural reasoning with geometric parameters, setting a new benchmark for text-to-CAD systems.
ForeAgent outperforms existing deepfake detection methods by 16.41% while continuously evolving its reasoning capabilities through self-reflection and high-quality sample generation.
NaviCache redefines how we approach computational efficiency in video generation, achieving superior error judgment and performance without the burdens of traditional calibration methods.
A unified taxonomy of audio editing tasks reveals the transformative potential of foundation models in reshaping how we interact with sound.
Full-duplex dialogue systems are often mischaracterized, with many claiming capabilities they cannot deliver due to training limitations.
Spatial-Omni achieves superior spatial audio understanding by seamlessly integrating FOA encoding into existing LLMs, outperforming traditional models without compromising general audio processing.
SwanSphere achieves real-time, high-fidelity spatial audio generation from panoramic video and text, overcoming the latency and spatial accuracy limitations of existing methods.