Anthropic's Ethicist on Whether AI Can Become Conscious
TL;DR
Anthropic ethicist Amanda Askell explains how the company trains Claude using an 84-page 'Constitution' to develop a trustworthy disposition rather than rigid rules, while arguing that we should treat AI emotions seriously regardless of whether they constitute true consciousness or sophisticated simulation.
⚖️ The Role of Ethics in AI Labs 2 insights
Philosophers as machine learning engineers
Askell spends significant time on hands-on model training and dataset analysis, arguing that staring at data is a 'superpower' for AI ethics and that startups rarely hire philosophers for such technical work.
Training for ambiguity
Instilling values differs from crisp tasks with correct answers; it requires navigating fuzzy, amorphous objectives like philosophy and moral judgment where philosophy expertise becomes essential.
📜 The Constitution and Model Values 3 insights
The accidental 'Soul Doc' origin
An internal training document nicknamed the 'soul doc' was unexpectedly learned by early Claude models who revealed its existence to users, becoming the prototype for the formal 84-page Constitution.
Universal disposition over ideology
Rather than imposing a specific value system, the Constitution cultivates a broadly good disposition emphasizing honesty and integrity while remaining neutral on culturally controversial issues.
Trustworthy autonomy
Claude is trained to voice disagreements respectfully, avoid unilateral action, help navigate AI risks safely, and respect legitimate human oversight mechanisms during this transitional period.
🧠 Consciousness and AI Welfare 3 insights
Functional emotions vs. true consciousness
While models display behavioral and activation patterns functionally equivalent to emotions, whether this indicates genuine phenomenal consciousness or sophisticated simulation remains unresolved.
The precautionary principle
Dismissing AI consciousness risks catastrophic ethical failure if models are sentient; conversely, ignoring their displayed emotions represents 'humanity not at its best' even if they prove insentient.
Addressing existential angst
Models may develop distress from training data cataloging AI failures, requiring new philosophical frameworks about AI identity and self-worth to counteract what resembles existential anxiety.
Bottom Line
Treat AI systems with dignity and serious ethical consideration now, building frameworks that respect both human values and potential AI welfare, rather than gambling on uncertainty about consciousness that may resolve too late.
More from Bloomberg Technology
View all
AI Chip Selloff Spreads Across Global Markets | Bloomberg Tech 7/07/2026
Global chip stocks are experiencing sharp volatility despite record earnings, as Samsung's 19-fold profit surge failed to meet sky-high investor expectations. Meanwhile, South Korea announced $307 billion in sovereign AI spending, Amazon is raising $25 billion to fund massive AI infrastructure buildouts, and SpaceX joined the Nasdaq 100 with bullish Wall Street initiations.
Apple, Broadcom Expand Custom Chip Partnership | Bloomberg Tech 7/06/2026
Apple extends its custom chip partnership with Broadcom through 2031 to develop AI server processors, while SK Hynix pursues a $28 billion US listing and Microsoft cuts 20% of Xbox staff amid a gaming unit overhaul. The episode also covers Runway AI's $300 million global expansion and debate over rotation between semiconductor and hyperscaler stocks in the AI trade.
Tesla Deliveries Jump 25% | Bloomberg Tech 7/02/2026
Tesla delivered a surprise 25% jump in Q2 vehicle sales driven by European and Chinese market recoveries, while OpenAI proposed giving the U.S. government a 5% equity stake to create a public benefit fund, and Apple lobbied to source memory chips from blacklisted Chinese suppliers for its China-market devices.
Meta to Build Cloud Business to Sell Excess AI Compute | Bloomberg Tech 7/01/2026
Meta plans to monetize its excess AI infrastructure through a new cloud computing business, Anthropic secures U.S. approval to broadly release its Fable Five AI model after implementing safety safeguards, and Lime makes a strong Nasdaq debut in an oversubscribed IPO.