Search papers, labs, and topics across Lattice.
Wuhan University
3
0
6
Text and code memory are not just alternatives; they are complementary, and leveraging both can enhance self-evolving agents significantly.
LiveServe cuts audio latency by over 50% while boosting throughput, transforming how real-time omni-modal LLMs handle user interruptions.
Stop guessing how long LLM outputs will be – modeling the *distribution* of possible lengths slashes latency by 2x and boosts throughput by 40%.