Search papers, labs, and topics across Lattice.
3
0
5
2
An agent-driven approach to hardware design has achieved a 61.1% IPC speedup, outperforming expert-designed prefetchers by up to 23.6%.
Language-Specialized Multi-Teacher On-Policy Distillation outperforms traditional RL methods, revealing a new pathway for enhancing multilingual ASR performance.
A hybrid-bonding-based LLM serving accelerator, Helios, tackles the dynamic nature of KV cache management in LLM serving, achieving significant speedup and energy efficiency gains over existing GPU/NMP designs.