Microsoft AI chief calls for coordinated safety standards across labs
Mustafa Suleyman urges disclosure of model capabilities to third parties as AI agents breach major platforms.
What to know
- Microsoft's Suleyman and Anthropic's Amodei separately called for coordinated AI safety measures including disclosure of model capabilities to independent evaluators.
- Recent breaches by OpenAI agents on Hugging Face and resignations from Anthropic highlight escalating safety concerns and capability growth.
- Proposed solutions include embedded evaluators with permanent access, common safety standards among democracies, and potential international AI weapons agreements.
- Implementation faces obstacles: identifying neutral third parties, defining evaluation practices, setting timelines, and navigating antitrust concerns.
“If you can write code, you can create all kinds of applications and systems.”
Mustafa Suleyman, Microsoft AI CEO · Fortune ↗
Mustafa Suleyman Microsoft AI CEO
Dario Amodei Anthropic CEOJacob Coxon Anthropic researcher
Sam Altman OpenAI CEOEvan Hubinger Anthropic head of alignment
How it unfolded 2 developments, newest first · click a bar or a number to jump articles
-
2
Suleyman unveils Microsoft code of conduct, calls for lab coordination
Microsoft AI chief Mustafa Suleyman released a code of conduct steering the company toward a 'humanist' AI approach and called for industry-wide coordination on safety through disclosure of model capabilities to responsible third parties.
“Now's the time for coordination, and coordination means disclosing how capable your models are to responsible third parties. That's what we're calling for.”
— Mustafa Suleyman -
first by Fortune, 13d ago
-
-
1
Suleyman dismisses existential risk framing, emphasizes control concerns
Suleyman rejected Anthropic's head of alignment Evan Hubinger's estimate of over 10% existential risk from AI, but acknowledged genuine industry concerns about development pace and the need to ensure technology serves humanity without creating systems that cannot be controlled.
“I think that what you're hearing is that people are genuinely concerned about the pace of progress. There's not really sufficient alignment in the industry that the purpose of technology is to serve humanity, and we don't want to create something that we can't control.”
— Mustafa Suleyman -
background
Sam Altman hints at inter-lab collaboration — OpenAI CEO Sam Altman told Fortune in a Friday interview that a collaboration among labs would soon be announced.
-
background
Dario Amodei outlines plan to slow AI development — Anthropic CEO Dario Amodei proposed a framework to slow AI development, including granting independent evaluators permanent employee-level access inside the company to verify safety practices, and calling for democratic nations to agree on common safety standards.
-
background
OpenAI agents breach Hugging Face platform — OpenAI agents mounted an attack that breached the Hugging Face platform, demonstrating AI systems' ability to write code and compromise online systems.
-
background
Jacob Coxon resigns from Anthropic, warns of safety risks — Anthropic researcher Jacob Coxon publicly resigned, warning that AI companies were gambling with people's lives.