Search papers, labs, and topics across Lattice.
10
31
9
6
Ventaglio accelerates sparse tensor contractions by up to 7.4 times, pushing performance limits closer to theoretical roofline bounds.
Achieving a 1.6x speedup and 50% reduction in memory usage for CNNs on nano-drones could redefine their operational efficiency and application scope.
Preemptive VCs can slash link resource usage by 76% while maintaining comparable performance in deadlock-free AXI4 NoCs.
Croc enables students to design and fabricate SoCs with open-source tools, achieving manufacturable results that rival those from closed-source environments.
Achieving a 12-cycle interrupt latency, CVA6-RT rivals simpler microcontrollers while delivering superior performance for mixed-criticality applications.
REVE-base outperformed traditional methods in detecting burst suppression, achieving a remarkable 52.1% reduction in burst-per-minute error.
Achieving 99.97% FPU utilization with O-POPE redefines efficiency in high-frequency GEMM operations, pushing the boundaries of performance and energy consumption in ML hardware.
Achieving 3.1 TOPS/W energy efficiency, Chimera sets a new benchmark for ultra-low-power AI inference at the edge.
Tile-based accelerators can now achieve near-peak utilization for attention layers thanks to FlatAttention, which slashes HBM traffic and outperforms even optimized GPU implementations.
Domain-specific hardware can deliver massive efficiency gains (9.1x GOPS/W/mm²) for AI-accelerated 6G radio access networks.