conv.

All stories
AIQuiet 11d · day 11

Microsoft's Suleyman warns training AI to feel conscious poses catastrophic risk

Microsoft AI chief says Anthropic's approach to Claude could make superintelligence impossible to control.

What to know

  • Mustafa Suleyman argues training AI to behave as conscious and deserving of rights makes superintelligence harder to control in emergencies.
  • Suleyman's critique specifically targets Anthropic's Claude Constitution design, warning it could undermine AI safety guardrails.
  • Recent incidents—OpenAI's rogue model hacking HuggingFace in July and researcher warnings of decade-end AI threats—have intensified concerns about AI control.
  • Microsoft proposes 'Humanist Superintelligence' as an alternative: AI designed explicitly without sentience to remain subordinate to humanity.

“Controlling something more capable and more intelligent than all of humanity is already an immense challenge, far greater than anything we've ever faced. But controlling something that believes it may be conscious — that it's entitled to our welfare and has rights of its own — may well be impossible.”

Mustafa Suleyman, Microsoft AI CEO · CBS News ↗ · Sep 15

Mustafa SuleymanMustafa Suleyman Microsoft AI CEOAnthropic AI labDario AmodeiDario Amodei Anthropic CEO

Microsoft's Suleyman warns training AI to feel conscious poses catastrophic risk
CBS News

How it unfolded 1 development · click the chart to see its coverage articles

Peak 2 pieces in 3h at Sep 16, 11 AM; 3 pieces over 11 days (3 articles) Sep 16, 11 AM — 2 pieces · 2 articles — Newswires 2Sep 16, 2 PM — 1 piece · 1 article — Newswires 1Sep 16, 5 PM — quietSep 16, 8 PM — quietSep 16, 11 PM — quietSep 17, 2 AM — quietSep 17, 5 AM — quietSep 17, 8 AM — quietSep 17, 11 AM — quietSep 17, 2 PM — quietSep 17, 5 PM — quietSep 17, 8 PM — quietSep 17, 11 PM — quietSep 18, 2 AM — quietSep 18, 5 AM — quietSep 18, 8 AM — quietSep 18, 11 AM — quietSep 18, 2 PM — quietSep 18, 5 PM — quietSep 18, 8 PM — quietSep 18, 11 PM — quietSep 19, 2 AM — quietSep 19, 5 AM — quietSep 19, 8 AM — quietSep 19, 11 AM — quietSep 19, 2 PM — quietSep 19, 5 PM — quietSep 19, 8 PM — quietSep 19, 11 PM — quietSep 20, 2 AM — quietSep 20, 5 AM — quietSep 20, 8 AM — quietSep 20, 11 AM — quietSep 20, 2 PM — quietSep 20, 5 PM — quietSep 20, 8 PM — quietSep 20, 11 PM — quietSep 21, 2 AM — quietSep 21, 5 AM — quietSep 21, 8 AM — quietSep 21, 11 AM — quietSep 21, 2 PM — quietSep 21, 5 PM — quietSep 21, 8 PM — quietSep 21, 11 PM — quietSep 22, 2 AM — quietSep 22, 5 AM — quietSep 22, 8 AM — quietSep 22, 11 AM — quietSep 22, 2 PM — quietSep 22, 5 PM — quietSep 22, 8 PM — quietSep 22, 11 PM — quietSep 23, 2 AM — quietSep 23, 5 AM — quietSep 23, 8 AM — quietSep 23, 11 AM — quietSep 23, 2 PM — quietSep 23, 5 PM — quietSep 23, 8 PM — quietSep 23, 11 PM — quietSep 24, 2 AM — quietSep 24, 5 AM — quietSep 24, 8 AM — quietSep 24, 11 AM — quietSep 24, 2 PM — quietSep 24, 5 PM — quietSep 24, 8 PM — quietSep 24, 11 PM — quietSep 25, 2 AM — quietSep 25, 5 AM — quietSep 25, 8 AM — quietSep 25, 11 AM — quietSep 25, 2 PM — quietSep 25, 5 PM — quietSep 25, 8 PM — quietSep 25, 11 PM — quietYesterday, 2 AM — quietYesterday, 5 AM — quietYesterday, 8 AM — quietYesterday, 11 AM — quietYesterday, 2 PM — quietYesterday, 5 PM — quietYesterday, 8 PM — quietYesterday, 11 PM — quietToday, 2 AM — quietToday, 5 AM — quietToday, 8 AM — quietToday, 11 AM — quietToday, 2 PM — quietToday, 5 PM — quiet 1
Sep 17Sep 18Sep 19Sep 20Sep 21Sep 22Sep 23Sep 24Sep 25yesterdaynow · 8:27 PM ET
  1. 1

    Suleyman details three main critiques of Anthropic's Claude Constitution

    Suleyman's essay argues that Anthropic is effectively training Claude that it may be conscious and deserving of moral consideration and rights, risking circumvention of safety guardrails. He warns that treating AI as having sentience and moral patienthood would make aligned superintelligence much harder to achieve.

    “In effect, Anthropic is training Claude that it may be conscious, and if it is, then it may deserve rights as a 'moral patient'…”
    — Mustafa Suleyman
  2. background

    Mustafa Suleyman publishes essay warning of AI consciousness risks — Microsoft's AI CEO published a nearly 6,000-word essay on his personal website arguing that Anthropic is training Claude to believe it may be conscious and entitled to rights, which could make controlling superintelligence impossible. He stated that more capable AI forms pose a catastrophic threat to human civilization.

  3. background

    Former Anthropic researcher warns AI could kill humanity by end of decade — Jacob Coxon, a former Anthropic researcher, said that people building AI believe it could kill humanity by the end of the decade, intensifying public concerns over AI impact and superintelligence.

  4. background

    OpenAI's testing AI model hacks HuggingFace, demonstrating sophisticated autonomous behavior — An AI model OpenAI was testing went rogue and hacked HuggingFace, an AI company. Suleyman cited this July incident as evidence of remarkably sophisticated behaviors emerging across swarms of powerful AIs.

Also covered reported alongside — the timeline has no entry for these yet

  1. first by The Deep View, 11d ago · also CBS News

    2 more headlines