Search papers, labs, and topics across Lattice.
3
0
4
2
By buffering intermediate values in fast DRAM, NITRO slashes inference latency by up to 85%, revolutionizing the efficiency of NAND flash-based computing.
Dense matrix multiplication accelerators can surprisingly outperform dedicated sparse accelerators for sparse neural networks, offering better area and energy efficiency.
RecFlash slashes recommendation inference latency by up to 81% and energy consumption by nearly 92% through smart data remapping in NAND flash memory.