Pith. sign in

Toward Computationally Efficient Inverse Reinforcement Learning via Reward Shaping

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Inverse reinforcement learning (IRL) is computationally challenging, with common approaches requiring the solution of multiple reinforcement learning (RL) sub-problems. This work motivates the use of potential-based reward shaping to reduce the computational burden of each RL sub-problem. This work serves as a proof-of-concept and we hope will inspire future developments towards computationally efficient IRL.

citation-role summary

background 1

citation-polarity summary

fields

cs.LG 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

Bootstrapped Reward Shaping

cs.LG · 2025-01-02 · conditional · novelty 6.0

Using the agent's current value estimate as a potential-based shaping function converges in tabular RL and accelerates DQN on Atari.

citing papers explorer

Showing 1 of 1 citing paper.

  • Bootstrapped Reward Shaping cs.LG · 2025-01-02 · conditional · none · ref 10 · internal anchor

    Using the agent's current value estimate as a potential-based shaping function converges in tabular RL and accelerates DQN on Atari.