Pith. sign in

REVIEW 1 cited by

Multi-Agent Reinforcement Learning for Power Grid Topology Optimization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.02605 v1 pith:I2ONBYS7 submitted 2023-10-04 cs.LG cs.AIcs.SYeess.SYstat.ML

Multi-Agent Reinforcement Learning for Power Grid Topology Optimization

classification cs.LG cs.AIcs.SYeess.SYstat.ML
keywords learningnetworkspowerreinforcementactionagentsdifferentframework
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

Recent challenges in operating power networks arise from increasing energy demands and unpredictable renewable sources like wind and solar. While reinforcement learning (RL) shows promise in managing these networks, through topological actions like bus and line switching, efficiently handling large action spaces as networks grow is crucial. This paper presents a hierarchical multi-agent reinforcement learning (MARL) framework tailored for these expansive action spaces, leveraging the power grid's inherent hierarchical nature. Experimental results indicate the MARL framework's competitive performance with single-agent RL methods. We also compare different RL algorithms for lower-level agents alongside different policies for higher-order agents.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Power Grid Control with Graph-Based Distributed Reinforcement Learning

    cs.LG 2025-09 conditional novelty 6.0

    A two-layer distributed RL system with one GNN-observing agent per power line and a learned manager keeps the Grid2Op case14 grid alive far longer than the do-nothing baseline.