Pith. sign in

Learning a belief representation for delayed reinforcement learning

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2024 1

verdicts

REJECT 1

representative citing papers

Inverse Delayed Reinforcement Learning

cs.LG · 2024-12-04 · reject · novelty 5.0

The paper proposes an off-policy adversarial IRL method on augmented delayed states and claims, via a Lipschitz bound and MuJoCo experiments, that this outperforms IRL on raw delayed observations.

citing papers explorer

Showing 1 of 1 citing paper.

  • Inverse Delayed Reinforcement Learning cs.LG · 2024-12-04 · reject · none · ref 8

    The paper proposes an off-policy adversarial IRL method on augmented delayed states and claims, via a Lipschitz bound and MuJoCo experiments, that this outperforms IRL on raw delayed observations.