Search papers, labs, and topics across Lattice.
3
0
6
5
ETHEREAL achieves sub-30 渭s latency for high-resolution vision tasks, setting a new standard for event-driven processing at the edge.
LUT-based hardware architectures can achieve up to 2.2x area reduction for LLM inference by challenging conventional design assumptions and optimizing for activation data types.
Mamba's quest for hyperscale GPU saturation has backfired, making it significantly slower on edge devices despite its theoretical efficiency.