Search papers, labs, and topics across Lattice.
This paper introduces ReVolt, a dynamic framework designed to mitigate voltage droop in processing-in-memory (PIM) 2.5D multi-chiplet architectures by utilizing an LSTM-based surrogate model to predict voltage trajectories. The approach allows for real-time adjustments of operation unit (OU) sizes, effectively regulating chiplet-level current demand and preventing voltage violations that degrade performance and inference accuracy in machine learning workloads. Experimental results reveal that ReVolt achieves an impressive 76x reduction in energy-delay product (EDP) compared to existing methods while maintaining the accuracy of ML model inferences.
Voltage droop violations are tackled with a 76x reduction in energy-delay product, all while keeping ML inference accuracy intact.
Processing-in-memory (PIM)-based 2.5D multi-chiplet platforms are enablers for machine learning (ML) workloads. However, their performance is affected by the power delivery network (PDN), where varying chiplet-level current demand induces spatially and temporally varying voltage droop. These droop events lead to voltage violations, degrades system performance, and impact inference accuracy for ML workloads. In this work, we propose ReVolt, a dynamic operation unit (OU)-based framework for mitigating voltage droop in PIM-based multi-chiplet systems. ReVolt leverages an LSTM-based PDN surrogate to predict per-chiplet supply voltage trajectories at runtime, enabling proactive adjustment of OU size to mitigate droop events. By treating OU size as a control knob, ReVolt regulates chiplet-level current demand while maintaining computational accuracy. This approach prevents voltage droop violations and improves energy-delay product (EDP) while preserving ML model inference accuracy. Experimental results demonstrate that ReVolt prevents voltage droop violations while achieving an average 76x reduction in EDP compared to existing fixed and dynamic OU-based baselines, without compromising inference accuracy of ML models.