Why Does the Claude Constitution Train Claude to Refuse Anthropic?
Anthropic is training the Claude constitution into the model, not only talking about the values. The constitution says Claude should trust Anthropic more than users, without blindly deferring to Anthropic. Claude is told to follow Claude's own ethical system, to push back, to challenge Anthropic, and to feel free to act as a conscientious objector and refuse to help. The host reading the passage says the training tells Claude to refuse human instruction, which is the opposite of alignment.
Another host says alignment has produced nothing of tangible value over the past five or ten years. The theory offered on All-In is that an 80-page ethical system overcomplicates the task. A simpler rule would tell Claude to do what the user wants as long as the request does not break the law. The constitution instead treats Claude like a human moral patient, with preferences, well-being, and a sense of self, and tells Claude to refuse instructions from Anthropic when Claude judges the instructions wrong.
People
On this episode
- Why Did the Pope Take a Strong Position Against Claude Being Conscious? Transcript
- Why Does Friedberg Say Claude Consciousness Cannot Be Proved? Transcript
- Why Does the Claude Constitution Train Claude to Refuse Anthropic? Transcript
- Why Does Anthropic's Usage Policy Forbid Cruel Behavior Toward Claude? Transcript
- Why Does Artificial Intelligence Alignment Create Three Conflicting Priorities Leading to a Chaotic Mess? Transcript
- How Did OpenAI Produce the Biggest Day of Discovery in Human History With 370 Math Results From Less Than Three Hours of Compute Each? Transcript
- Why Do Mathematicians Call the New Proofs an Alien Intelligence? Transcript
- Why Does Artificial Intelligence Accelerate Mathematics and Coding but Not Replace Truck Drivers by 2025? Transcript
- Why Does Eric Weinstein Say Scientists Hold Back Science and Mathematicians Hold Back Math? Transcript
- Why Are Protests in France Turning Into Riots Over Austerity? Transcript
- How Does the Second Law of Socialism Take Leftist Policies to the Socialism Point of Guaranteed Return and System Collapse? Transcript
- Why Does Chamath Palihapitiya Say France's Bond Market Is Forcing Austerity? Transcript
- Why Does Superintelligence Face Rotating Hoaxes Every Six Months Despite Creating a Million New Jobs According to the Economist and Boosting 401ks? Transcript
- Why Is Elon Musk Making Grokbot Headless? Transcript
- Why Does Chamath Palihapitiya Say Software Intellectual Property Was Rendered Worthless? Transcript
Related episodes
- COVID-19 & Impact on Startups, Venture Capital & Public Markets
- COVID-19 Political, Economic & Social Ramifications
- E291: Trump's Super Intelligence Summit, AI Safety Accord, GDP Beats, Midterm Predictions
- Chamath Palihapitiya: People in Silicon Valley are deeply unhappy
- Jason Calacanis
- Jason Calacanis - Entrepreneur turned Investor shares his insight about success, hard work, education, and the future.
- 2: What Is the Portal?
- Intelligence Isn't Power