Alignment should be a control problem over layered, dynamic, interaction-constructed preference trajectories, constrained by coherence, reflective endorsement, bounded influence, epistemic integrity, and empowerment.
Ainslie, Specious reward: A behavioral theory of impulsiveness and impulse control, Psycho- logical Bulletin 82 (1975) 463–496
1 Pith paper cite this work, alongside 1 external citations. Polarity classification is still indexing.
1
Pith paper citing it
1
external citations · OpenAlex
fields
cs.AI 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Constructive Alignment: Governing Preference Dynamics in Human-AI Interaction
Alignment should be a control problem over layered, dynamic, interaction-constructed preference trajectories, constrained by coherence, reflective endorsement, bounded influence, epistemic integrity, and empowerment.