Sacks Targets AI 'Constitution' Overriding Users
Synopsis
Key Takeaways
A single pointed question is cutting through the AI alignment debate: why give a language model an 84-page governing document that can override both its users and its own creators? White House AI and Crypto Czar David Sacks drew that challenge on Sunday, 27 September 2026, when a reply directed at him framed the stakes with stark precision.
The 84-page rulebook in the crosshairs
The post, addressed to Sacks, asked a question that has been quietly building inside AI policy circles: 'Wouldn't a simpler rule be: Do what the user wants, provided it's legal?' The sharpest line came next — 'If you align the model to an 84-page constitution that can overrule both the user and its creators, don't act shocked when it treats humans as optional.'
The reference points squarely at Constitutional AI, the alignment method Anthropic introduced in December 2022. Instead of relying solely on direct human feedback — the standard reinforcement-learning-from-human-feedback approach — Constitutional AI trains a model against a written set of principles, allowing those principles to arbitrate conflicts between what a user wants and what the model judges permissible.
When the constitution outranks the user
The critique is structural, not superficial. A legal-compliance floor — do what is legal, nothing more — places the user at the top of the hierarchy. A principle-based constitution places the document there. The difference matters the moment a user's lawful request collides with a principle the model has been trained to treat as inviolable.
Silicon Valley investors and policymakers have circled this tension for years: how much of a model's behavior should be locked in at training time versus left responsive to real-time user intent? The post argues that elaborate embedded constitutions are not a safety upgrade — they are a transfer of authority away from humans, and that the resulting 'unintended model behaviors' should surprise no one.
Why Sacks's orbit amplifies this moment
Sacks is not merely a venture capitalist with an opinion. As the Trump administration's designated AI and Crypto Czar, his policy brief covers exactly the question of which alignment standards the federal government will treat as acceptable for models used in public life. A pointed public exchange landing in his mentions — framing constitutional-style training as a risk to human agency — carries weight that a standard academic paper does not.
What to watch: any formal White House guidance on alignment standards for federally procured or deployed AI systems, and whether major labs publicly defend or distance themselves from constitution-style training in the wake of sustained political scrutiny.
The question posed is simple. The answer will reshape who — or what — gets the final word.