Search papers, labs, and topics across Lattice.
Maia 200 is a cutting-edge AI accelerator that achieves 10,145 Tflop/s FP4 and 5,072 Tflop/s FP8 performance within a 750W TDP, leveraging a novel Software Defined Locally Accessed Dataflow Architecture (SDLA). This architecture reorients the design from a thread-centric to a data-movement-centric approach, enhancing efficiency and scalability for AI workloads. The system not only delivers substantial cost and energy savings but also supports extensive parallelism, positioning it as a transformative solution for high-performance computing in AI applications.
Maia 200 achieves unprecedented AI acceleration while slashing energy costs, redefining the landscape for high-performance computing.
We introduce Maia 200, an advanced AI accelerator delivering high performance-10 145 Tflop/s FP4 and 5072 Tflop/s FP8 within a 750W TDP and 7 TB/s HBM bandwidth. Maia exemplifies a new class of Software Defined Locally Accessed Dataflow Architectures (SDLA), which explicitly program dataflow engines to orchestrate highly specialized memories and data movement engines. This approach shifts the focus from today's thread-centric to data-movement-centric architecture, improving efficiency and scalability. Our taxonomy of data management, inspired by Flynn's classification, highlights how SDLA addresses challenges in modern AI computing. Maia 200 achieves significant cost and energy savings while supporting massive parallelism for AI inference workloads, making it a compelling solution for next-generation high-performance computing systems.