The authors argue that RL agents should be classified as moral only relative to a distinct moral reward function, and use this criterion to assess three RL-based value alignment proposals.
Ethics118(1), 109–139 (2007)
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Metanormative Theory for RL-Based Moral Agents
The authors argue that RL agents should be classified as moral only relative to a distinct moral reward function, and use this criterion to assess three RL-based value alignment proposals.