Pith. sign in

Safe Reinforcement Learning by Imagining the Near Future

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Safe reinforcement learning is a promising path toward applying reinforcement learning algorithms to real-world problems, where suboptimal behaviors may lead to actual negative consequences. In this work, we focus on the setting where unsafe states can be avoided by planning ahead a short time into the future. In this setting, a model-based agent with a sufficiently accurate model can avoid unsafe states. We devise a model-based algorithm that heavily penalizes unsafe trajectories, and derive guarantees that our algorithm can avoid unsafe states under certain assumptions. Experiments demonstrate that our algorithm can achieve competitive rewards with fewer safety violations in several continuous control tasks.

citation-role summary

background 1

citation-polarity summary

fields

cs.RO 1

years

2024 1

verdicts

REJECT 1

roles

background 1

polarities

unclear 1

representative citing papers

Q-learning-based Model-free Safety Filter

cs.RO · 2024-11-29 · reject · novelty 6.0

A Q-learning safety filter with a time-dependent reward blocks unsafe actions from arbitrary task policies, but its theoretical guarantee is not valid as written.

citing papers explorer

Showing 1 of 1 citing paper.

  • Q-learning-based Model-free Safety Filter cs.RO · 2024-11-29 · reject · none · ref 23 · internal anchor

    A Q-learning safety filter with a time-dependent reward blocks unsafe actions from arbitrary task policies, but its theoretical guarantee is not valid as written.