Search papers, labs, and topics across Lattice.
This paper surveys nearly 800 studies at the intersection of graph machine learning (GML) and power systems, highlighting the potential of GML to address the operational complexities posed by renewable energy integration and real-time decision-making. It emphasizes the unique challenges of power systems, such as hard physical constraints and safety-critical requirements, which make them an ideal benchmark for GML applications. The authors also identify significant gaps in real-world deployment and reproducibility, advocating for standardized benchmarks and open datasets to enhance the credibility and impact of future research in this area.
Power systems could become the gold standard for testing graph machine learning, yet face a reproducibility crisis due to scarce benchmarks and datasets.
Modern power systems face growing operational complexity driven by the integration of renewable energy sources, decentralization, and the need for real-time decision-making across a wide range of timescales. Addressing these challenges traditionally relies on model-based methods that, while accurate, can be too slow for operational demands. Machine learning (ML) has therefore emerged as a faster, data-driven alternative. As grid topology plays a central role in power system operation, graph machine learning (GML) methods offer a natural framework for incorporating topological dependencies as an inductive bias. We survey nearly 800 papers at the intersection of GML and power systems, covering forecasting, state estimation, optimization, control, fault diagnosis, and cybersecurity. Power systems constitute an unusually rich benchmark setting for GML, as they combine hard physical constraints, multi-scale dynamics, safety-critical requirements, and scarce labeled data within a single, well-defined domain. Conversely, power systems can benefit from utilizing GML to complement classical solvers, as GML provide scalable, topology-aware approximations with promising generalization and computational efficiency. We identify open challenges, including limited real-world deployment and the need for interpretable models in safety-critical settings. Despite the rapidly growing number of publications, standardized benchmarks and open datasets remain scarce, leaving many results difficult to reproduce and undermining the long-term scientific credibility of the field. We further derive a structured requirements catalog for ML-ready power grid benchmarks, intended to guide future dataset development and improve reproducibility across studies. We call on the community to prioritize dedicated benchmark studies and the release of open datasets and models.