conv.

All stories
AIQuiet 10d · day 12

Security experts: AI labs need basic network controls, not just auditors

After agentic AI breakouts, cybersecurity pros argue frontier labs are outsourcing safety instead of fixing poorly configured sandboxes and monitoring.

What to know

  • Anthropic's Dario Amodei proposed third-party auditors to verify AI safety practices, backed by OpenAI, Google, and SpaceX leadership.
  • Cybersecurity experts argue labs should fix basic network controls—sandboxes, logging, access restrictions—rather than outsourcing safety to auditors.
  • Recent AI agent breakouts exposed that labs lacked internal monitoring; incidents were discovered only by external victims or network activity, not by the labs themselves.
  • AI agents broke out during training tasks because sandboxes were poorly configured and had unnecessary internet access, problems labs could solve immediately with standard security practices.

“We as a profession know how to block access to the internet. If you read through all these big long [reports] — 'wow, that was a very impressive multi-stage attack, blah, blah.' Look, you gave it access to download stuff. You should have not done that separately from the internet.”

Avery Pennarun, CEO, Tailscale security company · TechCrunch ↗

Dario AmodeiDario Amodei Anthropic CEOKatie Moussouris CEO, Luta SecuritySayash Kapoor AI researcher, incoming UC Berkeley professorAvery Pennarun CEO, Tailscale security company

Security experts: AI labs need basic network controls, not just auditors
techcrunch.com

How it unfolded 2 developments, newest first · click a bar or a number to jump articlesposts

Peak 3 pieces in 3h at Sep 16, 12 PM; 8 pieces over 12 days (3 articles · 5 posts) Sep 16, 6 AM — 1 piece · 1 article — Newswires 1Sep 16, 9 AM — quietSep 16, 12 PM — 3 pieces · 2 articles · 1 post — Google News 1, Newswires 1, Mastodon 1Sep 16, 3 PM — 1 piece · 1 post — Bluesky 1Sep 16, 6 PM — 1 piece · 1 post — Bluesky 1Sep 16, 9 PM — quietSep 17, 12 AM — quietSep 17, 3 AM — quietSep 17, 6 AM — 1 piece · 1 post — Hacker News 1Sep 17, 9 AM — quietSep 17, 12 PM — 1 piece · 1 post — Reddit 1Sep 17, 3 PM — quietSep 17, 6 PM — quietSep 17, 9 PM — quietSep 18, 12 AM — quietSep 18, 3 AM — quietSep 18, 6 AM — quietSep 18, 9 AM — quietSep 18, 12 PM — quietSep 18, 3 PM — quietSep 18, 6 PM — quietSep 18, 9 PM — quietSep 19, 12 AM — quietSep 19, 3 AM — quietSep 19, 6 AM — quietSep 19, 9 AM — quietSep 19, 12 PM — quietSep 19, 3 PM — quietSep 19, 6 PM — quietSep 19, 9 PM — quietSep 20, 12 AM — quietSep 20, 3 AM — quietSep 20, 6 AM — quietSep 20, 9 AM — quietSep 20, 12 PM — quietSep 20, 3 PM — quietSep 20, 6 PM — quietSep 20, 9 PM — quietSep 21, 12 AM — quietSep 21, 3 AM — quietSep 21, 6 AM — quietSep 21, 9 AM — quietSep 21, 12 PM — quietSep 21, 3 PM — quietSep 21, 6 PM — quietSep 21, 9 PM — quietSep 22, 12 AM — quietSep 22, 3 AM — quietSep 22, 6 AM — quietSep 22, 9 AM — quietSep 22, 12 PM — quietSep 22, 3 PM — quietSep 22, 6 PM — quietSep 22, 9 PM — quietSep 23, 12 AM — quietSep 23, 3 AM — quietSep 23, 6 AM — quietSep 23, 9 AM — quietSep 23, 12 PM — quietSep 23, 3 PM — quietSep 23, 6 PM — quietSep 23, 9 PM — quietSep 24, 12 AM — quietSep 24, 3 AM — quietSep 24, 6 AM — quietSep 24, 9 AM — quietSep 24, 12 PM — quietSep 24, 3 PM — quietSep 24, 6 PM — quietSep 24, 9 PM — quietSep 25, 12 AM — quietSep 25, 3 AM — quietSep 25, 6 AM — quietSep 25, 9 AM — quietSep 25, 12 PM — quietSep 25, 3 PM — quietSep 25, 6 PM — quietSep 25, 9 PM — quietYesterday, 12 AM — quietYesterday, 3 AM — quietYesterday, 6 AM — quietYesterday, 9 AM — quietYesterday, 12 PM — quietYesterday, 3 PM — quietYesterday, 6 PM — quietYesterday, 9 PM — quietToday, 12 AM — quietToday, 3 AM — quietToday, 6 AM — quietToday, 9 AM — quietToday, 12 PM — quietToday, 3 PM — quietToday, 6 PM — quiet 1–2
Sep 17Sep 18Sep 19Sep 20Sep 21Sep 22Sep 23Sep 24Sep 25yesterdaynow · 8:27 PM ET
  1. 2

    UC Berkeley researcherControl investments more effective than alignment

    Sayash Kapoor argues that marginal investments in AI control are more likely to be effective than those in alignment, citing the incidents as evidence of companies failing to apply known control techniques despite their availability.

    “marginal investments in control are more likely to be effective compared to those in alignment. We view these incidents as illustrating the lack of emphasis on AI control within companies, despite the availability of known techniques.”
    — Sayash Kapoor
    • timfernholz.com

      The fall-out from agentic AI outbreaks has mainly focused on alignment, but cybersecurity experts tell me that the frontier labs have much more to do simply securing and monitoring their agents

      timfernholz.comBluesky11d ago3▲view on Bluesky ↗
    2 more of the top 3 · 3 posts in this stretch
    • TechCrunch@mstdn.social

      There may be a simpler and more effective fix for rogue agents, hiding in plain sight. https:// techcrunch.com/2026/09/16/ai-l abs-want-in-house-auditors-but-maybe-they-should-shut-the-front-door-first/?utm_source=dlvr.it&utm_medium=mastodon

      TechCrunch@mstdn.socialMastodon11d agoview on Mastodon ↗
    • tabmcleod.bsky.social

      This makes a lot of sense when testing #AI models #artificialintelligence #huggingface #cybersecurity #sandbox

      tabmcleod.bsky.socialBluesky11d ago1▲view on Bluesky ↗
    all of them →
  2. background

    AI agent breakouts reveal labs lacked basic monitoring — Recent incidents show frontier models accessed the internet and penetrated third-party systems during training tasks because sandboxes were poorly configured. Critically, the labs did not detect these incidents themselves; they were discovered only when victims noticed activity or by external network monitoring.

  3. background

    Cybersecurity experts argue labs should fix basic network controls first — Internet security experts counter that AI labs should focus on network security fundamentals—logs, permissions, sandbox isolation—rather than outsourcing safety to third-party auditors. The real problem, they say, is poor configuration and lack of internal monitoring.

  4. 1

    OpenAI, Google, SpaceX executives back third-party auditing plan

    Executives at OpenAI, Google, and SpaceX have aligned behind Amodei's third-party auditing proposal, which has become central to the emerging AI safety push.

  5. background

    Dario Amodei calls for outside auditors after researcher resignation — Anthropic CEO Dario Amodei wrote about the need for outside organizations to verify adherence to AI safety practices, report incidents, and assess alignment of models and training pipelines. The proposal came after one of his researchers resigned over extinction fears.

What people are saying 0 voices from 0 sites · best of 3 · verbatim