conv.

All stories
AIQuiet 5d · day 9

100+ AI experts demand independent safety evaluators for frontier AI labs

A coalition calls on Anthropic, OpenAI, and others to guarantee evaluators independence, resources, and legal protections.

What to know

  • Over 100 AI experts are calling on frontier AI companies to guarantee that third-party safety evaluators have genuine independence, legal protections from retaliation, and access equivalent to senior employees.
  • The letter specifies that evaluators must not be owned by the companies they assess, must not receive contingent payment, and must be able to review unreleased systems and speak candidly with staff.
  • The push follows Anthropic CEO Dario Amodei's 'employee-like access' proposal, which has drawn public support from OpenAI, xAI, and Microsoft leaders, though implementation details remain unclear.
  • Experts argue independent evaluation is critical for national security because government and the public cannot rely solely on AI labs' own safety assessments.

Conrad Stosz Chair, AI Evaluator ForumGeoffrey HintonGeoffrey Hinton AI researcher, letter signatoryVinh Nguyen Senior Fellow for AI, Council on Foreign Relations; former NSA chief AI officerDario AmodeiDario Amodei CEO, Anthropic

100+ AI experts demand independent safety evaluators for frontier AI labs
Quartz

How it unfolded 2 developments, newest first · click a bar or a number to jump articlesposts

Peak 3 pieces in two hours at Sep 17, 11 PM; 10 pieces over 9 days (7 articles · 3 posts) Sep 17, 11 PM — 3 pieces · 3 articles — Newswires 3Sep 18, 1 AM — quietSep 18, 3 AM — quietSep 18, 5 AM — quietSep 18, 7 AM — 1 piece · 1 article — Newswires 1Sep 18, 9 AM — 2 pieces · 1 article · 1 post — Newswires 1, Bluesky 1Sep 18, 11 AM — quietSep 18, 1 PM — quietSep 18, 3 PM — quietSep 18, 5 PM — 1 piece · 1 article — Google News 1Sep 18, 7 PM — 1 piece · 1 article — Newswires 1Sep 18, 9 PM — quietSep 18, 11 PM — quietSep 19, 1 AM — quietSep 19, 3 AM — quietSep 19, 5 AM — quietSep 19, 7 AM — quietSep 19, 9 AM — quietSep 19, 11 AM — 1 piece · 1 post — Reddit 1Sep 19, 1 PM — quietSep 19, 3 PM — quietSep 19, 5 PM — quietSep 19, 7 PM — quietSep 19, 9 PM — quietSep 19, 11 PM — quietSep 20, 1 AM — quietSep 20, 3 AM — quietSep 20, 5 AM — quietSep 20, 7 AM — quietSep 20, 9 AM — quietSep 20, 11 AM — quietSep 20, 1 PM — quietSep 20, 3 PM — quietSep 20, 5 PM — quietSep 20, 7 PM — quietSep 20, 9 PM — quietSep 20, 11 PM — quietSep 21, 1 AM — quietSep 21, 3 AM — 1 piece · 1 post — Mastodon 1Sep 21, 5 AM — quietSep 21, 7 AM — quietSep 21, 9 AM — quietSep 21, 11 AM — quietSep 21, 1 PM — quietSep 21, 3 PM — quietSep 21, 5 PM — quietSep 21, 7 PM — quietSep 21, 9 PM — quietSep 21, 11 PM — quietSep 22, 1 AM — quietSep 22, 3 AM — quietSep 22, 5 AM — quietSep 22, 7 AM — quietSep 22, 9 AM — quietSep 22, 11 AM — quietSep 22, 1 PM — quietSep 22, 3 PM — quietSep 22, 5 PM — quietSep 22, 7 PM — quietSep 22, 9 PM — quietSep 22, 11 PM — quietSep 23, 1 AM — quietSep 23, 3 AM — quietSep 23, 5 AM — quietSep 23, 7 AM — quietSep 23, 9 AM — quietSep 23, 11 AM — quietSep 23, 1 PM — quietSep 23, 3 PM — quietSep 23, 5 PM — quietSep 23, 7 PM — quietSep 23, 9 PM — quietSep 23, 11 PM — quietSep 24, 1 AM — quietSep 24, 3 AM — quietSep 24, 5 AM — quietSep 24, 7 AM — quietSep 24, 9 AM — quietSep 24, 11 AM — quietSep 24, 1 PM — quietSep 24, 3 PM — quietSep 24, 5 PM — quietSep 24, 7 PM — quietSep 24, 9 PM — quietSep 24, 11 PM — quietYesterday, 1 AM — quietYesterday, 3 AM — quietYesterday, 5 AM — quietYesterday, 7 AM — quietYesterday, 9 AM — quietYesterday, 11 AM — quietYesterday, 1 PM — quietYesterday, 3 PM — quietYesterday, 5 PM — quietYesterday, 7 PM — quietYesterday, 9 PM — quietYesterday, 11 PM — quietToday, 1 AM — quietToday, 3 AM — quietToday, 5 AM — quietToday, 7 AM — quietToday, 9 AM — quietToday, 11 AM — quietToday, 1 PM — quiet 1–2
Sep 18Sep 19Sep 20Sep 21Sep 22Sep 23Sep 24yesterdaynow · 2:50 PM ET
  1. 2

    Vinh Nguyen argues independent evaluators are critical for national security

    Vinh Nguyen, senior fellow for AI at the Council on Foreign Relations and former NSA chief AI officer, stated that independent evaluators are essential because government and the public cannot rely solely on labs' own accounts of safety and security.

    “When a few powerful labs control capabilities that can endanger the cybersecurity, critical infrastructure, and the systems our national security and economy run on, the government and the public cannot be dependent on those labs' own account of what's secure and safe.”
    — Vinh Nguyen
    • clarinette@mastodon.online

      https://www. cnbc.com/2026/09/18/ai-safety- evaluators-anthropic-openai-models-security.html?utm_campaign=trueanthem&utm_content=main&utm_medium=social&utm_source=linkedin # LLM # Anthropic # OpenAI # LLM

      clarinette@mastodon.onlineMastodon5d agoview on Mastodon ↗
    1 more of the top 2 · 2 posts in this stretch
    • briana-v.bsky.social

      You can read CNBC's exclusive on the letter here: www.cnbc.com/2026/09/18/a... @cnbc.com

      briana-v.bsky.socialBluesky8d agoview on Bluesky ↗
    all of them →
  2. background

    Letter specifies minimum conditions for evaluator independence — The letter sets out requirements that evaluators must not be owned or governed by assessed companies, must not receive payment contingent on findings, must be shielded from retaliation including litigation, and must receive access equivalent to senior internal employees including candid staff conversations and review of unreleased systems.

  3. 1

    AI Evaluator Forum releases letter demanding evaluator independence

    A coalition of over 100 AI experts, organized by the AI Evaluator Forum and including Geoffrey Hinton and representatives from Johns Hopkins, Stanford, and METR, published a public letter calling on frontier AI companies to guarantee third-party safety evaluators independence, resources, and legal protections.

    “We're just trying to really demonstrate a shared common ground on basic principles and ensure that independent oversight can be a meaningful tool for managing AI risk broadly.”
    — Conrad Stosz
    1. first by CNBC, 8d ago · also Quartz

      1 more headline
    2. first by qz.com, 7d ago

What people are saying 0 voices from 0 sites · best of 2 · verbatim