Reward shaping penalties for low gas-storage levels and inactivity substantially improve deep reinforcement learning for long-horizon power-to-gas economic dispatch, though results are in-sample on one data year.
Inte- grated Electricity-Gas System Optimal Dispatch Based on Deep Reinforcement Learning
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
eess.SY 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
The Economic Dispatch of Power-to-Gas Systems with Deep Reinforcement Learning:Tackling the Challenge of Delayed Rewards with Long-Term Energy Storage
Reward shaping penalties for low gas-storage levels and inactivity substantially improve deep reinforcement learning for long-horizon power-to-gas economic dispatch, though results are in-sample on one data year.