Pith. sign in

Constrained policy optimization

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Skill-based Safe Reinforcement Learning with Risk Planning

cs.LG · 2025-05-02 · conditional · novelty 5.0

SSkP uses PU learning to build a skill risk predictor from demonstrations and a cross-entropy risk planner to select safe skills during online RL, outperforming prior methods on most tested MuJoCo tasks.

citing papers explorer

Showing 1 of 1 citing paper.

  • Skill-based Safe Reinforcement Learning with Risk Planning cs.LG · 2025-05-02 · conditional · none · ref 2

    SSkP uses PU learning to build a skill risk predictor from demonstrations and a cross-entropy risk planner to select safe skills during online RL, outperforming prior methods on most tested MuJoCo tasks.