No AI system should be developed that does not hold commonly agreed upon values.
Thesis
No Intentional Misalignment
Summary
The case for including this
Forbidding the development of AI that lacks commonly agreed values targets the core technical problem of alignment, insisting that capability never outpace our ability to ensure a system shares human ethical commitments. It encodes the intuition that the danger of advanced AI lies chiefly in misaligned goals, and makes value-alignment a precondition rather than an afterthought. As a guiding requirement it orients development toward safety from the outset.
The case for changing or excluding this
Including this trades value pluralism against alignment: 'commonly agreed upon values' papers over deep, persistent moral disagreement among cultures and individuals, leaving the standard either vacuously thin or imposing one group's values as universal. It is also not operationally testable, since we lack reliable means to verify a system's values, making compliance unfalsifiable. The clause should specify alignment processes and a thin floor of widely shared commitments rather than presuppose a consensus on values that does not exist.
Discussion
Sign in to join the discussion.
Related resources
- Collective Constitutional AI↗DiscussesCollective Intelligence Project (with Anthropic) · External Resource · Oct 17, 2023
- AI 2027↗SupportsKokotajlo, Lifland, Larsen, Dean & Alexander (AI Futures Project) · External Resource · Apr 3, 2025
- Machine Intelligence Research Institute · External Resource
- Alignment Assemblies↗DiscussesCollective Intelligence Project · External Resource
- Collective Intelligence Project · External Resource
- Anthropic · External Resource
- Nick Bostrom (Oxford University Press, 2014) · External Resource
- Democratic Inputs to AI↗DiscussesOpenAI · External Resource · May 25, 2023
- If Anyone Builds It, Everyone Dies↗SupportsWikipedia · External Resource
- Anthropic · External Resource
- Asilomar AI Principles↗SupportsFuture of Life Institute · External Resource
- The Techno-Optimist Manifesto↗ChallengesMarc Andreessen (a16z) · External Resource
- Future of Life Institute · External Resource
- UK Government (GOV.UK) · External Resource
- AI Snake Oil↗ChallengesNarayanan & Kapoor (Princeton University Press) · External Resource