Pith. sign in

Title resolution pending

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.AI 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Safe Planning and Policy Optimization via World Model Learning

cs.AI · 2025-06-05 · conditional · novelty 6.0

SPOWL is a model-based safe RL method that uses a value-equivalent world model, a Lagrangian-trained safe policy, and adaptive planning thresholds to achieve low-cost, high-reward control on SafetyGymnasium tasks.

citing papers explorer

Showing 1 of 1 citing paper.

  • Safe Planning and Policy Optimization via World Model Learning cs.AI · 2025-06-05 · conditional · none · ref 5

    SPOWL is a model-based safe RL method that uses a value-equivalent world model, a Lagrangian-trained safe policy, and adaptive planning thresholds to achieve low-cost, high-reward control on SafetyGymnasium tasks.