Pith. sign in

Foundations for restraining bolts: Reinforcement learning with ltlf/ldlf restraining specifications

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.RO 1

years

2024 1

verdicts

CONDITIONAL 1

representative citing papers

Adaptive Reward Design for Reinforcement Learning

cs.RO · 2024-12-14 · conditional · novelty 6.0

An adaptive reward-shaping method that periodically inflates distance-to-acceptance values for under-progressed stages lets RL agents reach the best achievable task progression on co-safe LTL tasks within a finite number of updates.

citing papers explorer

Showing 1 of 1 citing paper.

  • Adaptive Reward Design for Reinforcement Learning cs.RO · 2024-12-14 · conditional · none · ref 6

    An adaptive reward-shaping method that periodically inflates distance-to-acceptance values for under-progressed stages lets RL agents reach the best achievable task progression on co-safe LTL tasks within a finite number of updates.