REVIEW 6 cited by
RL2Grid: Benchmarking Reinforcement Learning in Power Grid Operations
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Reinforcement learning (RL) can provide adaptive and scalable controllers essential for power grid decarbonization. However, RL methods struggle with power grids' complex dynamics, long-horizon goals, and hard physical constraints. For these reasons, we present RL2Grid, a benchmark designed in collaboration with power system operators to accelerate progress in grid control and foster RL maturity. Built on RTE France's power simulation framework, RL2Grid standardizes tasks, state and action spaces, and reward structures for a systematic evaluation and comparison of RL algorithms. Moreover, we integrate operational heuristics and design safety constraints based on human expertise to ensure alignment with physical requirements. By establishing reference performance metrics for classic RL baselines on RL2Grid's tasks, we highlight the need for novel methods capable of handling real systems and discuss future directions for RL-based grid control.
Forward citations
Cited by 6 Pith papers
-
Robustness of Reinforcement Learning-Based Congestion Management in Low-Voltage Grids
A two-step random-forest-plus-actor-critic controller reduces congestion violations by 98.9% under accurate grid parameters, stays robust to measurement noise, but degrades to 79.6% under grid-model mismatch.
-
OpenG2G: A Simulation Platform for AI Datacenter-Grid Runtime Coordination
OpenG2G is a new extensible simulation platform that lets users implement and compare classic, optimization, and learning-based controllers for AI datacenter power flexibility coordinated with the grid.
-
MARS-DA: A Hierarchical Reinforcement Learning Framework for Risk-Aware Multi-Agent Bidding in Power Grids
MARS-DA uses a top-level meta-controller to blend safe day-ahead allocation and real-time arbitrage sub-policies, delivering better risk-adjusted returns than baselines in a PJM-grounded two-settlement market simulator.
-
Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift
A randomized audit certifies, with high confidence, that skipped N-1 contingencies violate thermal limits at most a chosen rate even under deployment shift, cutting full AC studies by 29–75% on three test systems.
-
Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning
RNN-ProVe uses policy-driven sampling and statistical error bounds to produce high-confidence probabilistic estimates of behavioral violations in RNN policies for single- and multi-agent POMDPs.
-
Interpretable Policy Distillation for Power Grid Topology Control
PPO policy for grid topology control is distilled into decision trees and random forests that outperform the teacher on reward and survival time with lower inference cost and high interpretability.
Discussion (0). Sign in to comment.