Standard inverse reinforcement learning models identify rewards only up to reward shaping and redistribution, and are not robust to even small misspecifications of transition dynamics or discount factors.
Definition 78 is simply directly analogous to Definition 7, except that it makes the assumption that the learning algorithm L uses the inductive bias described by L
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2024 1verdicts
ACCEPT 1representative citing papers
citing papers explorer
-
Partial Identifiability and Misspecification in Inverse Reinforcement Learning
Standard inverse reinforcement learning models identify rewards only up to reward shaping and redistribution, and are not robust to even small misspecifications of transition dynamics or discount factors.