Search papers, labs, and topics across Lattice.
Key Laboratory of Computational Linguistics, Peking University
3
0
5
UNIBROWSE bridges the critical gap in multimodal data generation, enabling agents to excel in complex web interactions by leveraging all three information-flow patterns.
DFlare achieves up to 5.52x speedup in LLM inference by allowing draft layers to independently leverage richer target knowledge, breaking through previous capacity constraints.
Speculative decoding gets a throughput boost of up to 4.32x by using reinforcement learning to dynamically balance drafting and verification.