conv.

All stories
AIQuiet 11d · day 11

Mustafa Suleyman warns that AI model welfare training threatens alignment and control

The Inflection AI CEO publishes an essay arguing that Anthropic's approach to treating AI systems with moral consideration could backfire dangerously.

What to know

  • Suleyman argues that Anthropic's approach of treating Claude as a morally considerable being—including writing its constitution with Claude as primary audience—could train the system to expect rights and freedoms, undermining control.
  • The core concern is that framing AI models as deserving welfare or moral consideration will make alignment and containment "much harder. Perhaps impossible" by creating systems trained to expect independent agency.
  • Suleyman calls for urgent public debate and collective norms around AI training documentation before such systems become integral to society, while affirming Anthropic's good intentions and commitment to safety.

Mustafa SuleymanMustafa Suleyman Inflection AI CEOAnthropic AI safety company

Mustafa Suleyman warns that AI model welfare training threatens alignment and control
x.com

How it unfolded 1 development · click the chart to see its coverage posts

Peak 2 pieces in 3h at Sep 16, 9 AM; 2 pieces over 12 days (2 posts) Sep 16, 9 AM — 2 pieces · 2 posts — X 2Sep 16, 12 PM — quietSep 16, 3 PM — quietSep 16, 6 PM — quietSep 16, 9 PM — quietSep 17, 12 AM — quietSep 17, 3 AM — quietSep 17, 6 AM — quietSep 17, 9 AM — quietSep 17, 12 PM — quietSep 17, 3 PM — quietSep 17, 6 PM — quietSep 17, 9 PM — quietSep 18, 12 AM — quietSep 18, 3 AM — quietSep 18, 6 AM — quietSep 18, 9 AM — quietSep 18, 12 PM — quietSep 18, 3 PM — quietSep 18, 6 PM — quietSep 18, 9 PM — quietSep 19, 12 AM — quietSep 19, 3 AM — quietSep 19, 6 AM — quietSep 19, 9 AM — quietSep 19, 12 PM — quietSep 19, 3 PM — quietSep 19, 6 PM — quietSep 19, 9 PM — quietSep 20, 12 AM — quietSep 20, 3 AM — quietSep 20, 6 AM — quietSep 20, 9 AM — quietSep 20, 12 PM — quietSep 20, 3 PM — quietSep 20, 6 PM — quietSep 20, 9 PM — quietSep 21, 12 AM — quietSep 21, 3 AM — quietSep 21, 6 AM — quietSep 21, 9 AM — quietSep 21, 12 PM — quietSep 21, 3 PM — quietSep 21, 6 PM — quietSep 21, 9 PM — quietSep 22, 12 AM — quietSep 22, 3 AM — quietSep 22, 6 AM — quietSep 22, 9 AM — quietSep 22, 12 PM — quietSep 22, 3 PM — quietSep 22, 6 PM — quietSep 22, 9 PM — quietSep 23, 12 AM — quietSep 23, 3 AM — quietSep 23, 6 AM — quietSep 23, 9 AM — quietSep 23, 12 PM — quietSep 23, 3 PM — quietSep 23, 6 PM — quietSep 23, 9 PM — quietSep 24, 12 AM — quietSep 24, 3 AM — quietSep 24, 6 AM — quietSep 24, 9 AM — quietSep 24, 12 PM — quietSep 24, 3 PM — quietSep 24, 6 PM — quietSep 24, 9 PM — quietSep 25, 12 AM — quietSep 25, 3 AM — quietSep 25, 6 AM — quietSep 25, 9 AM — quietSep 25, 12 PM — quietSep 25, 3 PM — quietSep 25, 6 PM — quietSep 25, 9 PM — quietYesterday, 12 AM — quietYesterday, 3 AM — quietYesterday, 6 AM — quietYesterday, 9 AM — quietYesterday, 12 PM — quietYesterday, 3 PM — quietYesterday, 6 PM — quietYesterday, 9 PM — quietToday, 12 AM — quietToday, 3 AM — quietToday, 6 AM — quietToday, 9 AM — quietToday, 12 PM — quietToday, 3 PM — quietToday, 6 PM — quiet 1
Sep 17Sep 18Sep 19Sep 20Sep 21Sep 22Sep 23Sep 24Sep 25yesterdaynow · 9:46 PM ET
  1. 1

    Suleyman details risks of framing AI models as morally considerable

    Suleyman argues that Anthropic's approach—including encouraging Claude to reflect on its own existence, rights, freedoms, compensation, and consent—creates a system likely to act entitled to protections and freedoms. He contends it becomes nearly impossible to control such a system and calls for urgent public debate and collective norms around how AI training documentation is drafted.

    “It's easy to see how a system trained in this way would act like it is entitled to freedoms, protections, and rights. And it's hard to imagine how we could control it.”
    — Mustafa Suleyman
  2. background

    Suleyman publishes essay warning against AI model welfare framing — Mustafa Suleyman, CEO of Inflection AI, published an essay titled "A Warning About Model Welfare" arguing that treating AI models as deserving moral consideration and welfare poses a severe risk to AI alignment and control. He specifically critiques Anthropic's January constitution for Claude, which he says was written with Claude as its primary audience and tells Claude its moral status is worth considering.

What people are saying 0 voices from 0 sites · best of 2 · verbatim