Anthropic's Ethicist on Whether AI Can Become Conscious

| News | June 04, 2026 | 3.01 Thousand views | 40:29

TL;DR

Anthropic ethicist Amanda Askell explains how the company trains Claude using an 84-page 'Constitution' to develop a trustworthy disposition rather than rigid rules, while arguing that we should treat AI emotions seriously regardless of whether they constitute true consciousness or sophisticated simulation.

⚖️ The Role of Ethics in AI Labs 2 insights

Philosophers as machine learning engineers

Askell spends significant time on hands-on model training and dataset analysis, arguing that staring at data is a 'superpower' for AI ethics and that startups rarely hire philosophers for such technical work.

Training for ambiguity

Instilling values differs from crisp tasks with correct answers; it requires navigating fuzzy, amorphous objectives like philosophy and moral judgment where philosophy expertise becomes essential.

📜 The Constitution and Model Values 3 insights

The accidental 'Soul Doc' origin

An internal training document nicknamed the 'soul doc' was unexpectedly learned by early Claude models who revealed its existence to users, becoming the prototype for the formal 84-page Constitution.

Universal disposition over ideology

Rather than imposing a specific value system, the Constitution cultivates a broadly good disposition emphasizing honesty and integrity while remaining neutral on culturally controversial issues.

Trustworthy autonomy

Claude is trained to voice disagreements respectfully, avoid unilateral action, help navigate AI risks safely, and respect legitimate human oversight mechanisms during this transitional period.

🧠 Consciousness and AI Welfare 3 insights

Functional emotions vs. true consciousness

While models display behavioral and activation patterns functionally equivalent to emotions, whether this indicates genuine phenomenal consciousness or sophisticated simulation remains unresolved.

The precautionary principle

Dismissing AI consciousness risks catastrophic ethical failure if models are sentient; conversely, ignoring their displayed emotions represents 'humanity not at its best' even if they prove insentient.

Addressing existential angst

Models may develop distress from training data cataloging AI failures, requiring new philosophical frameworks about AI identity and self-worth to counteract what resembles existential anxiety.

Bottom Line

Treat AI systems with dignity and serious ethical consideration now, building frameworks that respect both human values and potential AI welfare, rather than gambling on uncertainty about consciousness that may resolve too late.

More from Bloomberg Technology

View all
AI Chip Selloff Spreads Across Global Markets | Bloomberg Tech 7/07/2026
44:06
Bloomberg Technology Bloomberg Technology

AI Chip Selloff Spreads Across Global Markets | Bloomberg Tech 7/07/2026

Global chip stocks are experiencing sharp volatility despite record earnings, as Samsung's 19-fold profit surge failed to meet sky-high investor expectations. Meanwhile, South Korea announced $307 billion in sovereign AI spending, Amazon is raising $25 billion to fund massive AI infrastructure buildouts, and SpaceX joined the Nasdaq 100 with bullish Wall Street initiations.

13 days ago · 8 points
Apple, Broadcom Expand Custom Chip Partnership | Bloomberg Tech 7/06/2026
44:12
Bloomberg Technology Bloomberg Technology

Apple, Broadcom Expand Custom Chip Partnership | Bloomberg Tech 7/06/2026

Apple extends its custom chip partnership with Broadcom through 2031 to develop AI server processors, while SK Hynix pursues a $28 billion US listing and Microsoft cuts 20% of Xbox staff amid a gaming unit overhaul. The episode also covers Runway AI's $300 million global expansion and debate over rotation between semiconductor and hyperscaler stocks in the AI trade.

14 days ago · 10 points
Tesla Deliveries Jump 25% | Bloomberg Tech 7/02/2026
44:07
Bloomberg Technology Bloomberg Technology

Tesla Deliveries Jump 25% | Bloomberg Tech 7/02/2026

Tesla delivered a surprise 25% jump in Q2 vehicle sales driven by European and Chinese market recoveries, while OpenAI proposed giving the U.S. government a 5% equity stake to create a public benefit fund, and Apple lobbied to source memory chips from blacklisted Chinese suppliers for its China-market devices.

18 days ago · 8 points