The paper argues that aligning AGI to human goals does not automatically eliminate catastrophic risk, because alignment techniques can make powerful AI easier to misuse, and calls for safety research that avoids this tradeoff.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CY 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Misalignment or misuse? The AGI alignment tradeoff
The paper argues that aligning AGI to human goals does not automatically eliminate catastrophic risk, because alignment techniques can make powerful AI easier to misuse, and calls for safety research that avoids this tradeoff.