Part of The AI Control Crisis · 11 stories · since Sep 4 · newest 31m ago
OpenAI discloses six new AI ‘misalignment’ incidents, unveils reporting frameworkOnline reaction splits between alarm and accusations of PR spin
6 Sep 17 5:00 AM · 7d ago · 26 articles · 1 video · 25 posts · 8 comments · 7 sources · development 6 of 8
Commenters on Bluesky, Hacker News and Reddit debated whether the disclosures reflect genuine emergent risk or a self-serving narrative timed to influence AI regulation and funding.
“One of the key problems with AI is that the AI bros and their behemoth companies have absolutely no governance or guardrails. It's not that the AI is conscious or going rogue or whatever. It's that it's run by feckless arseholes.”
katebevan.comOpenAI AI developer disclosing the incidents
Sam Altman OpenAI CEO
Dario Amodei Anthropic CEOChris Lehane OpenAI global policy chief
David Sacks White House AI advisor
The whole story articlesvideospostscomments the bright band is this development · numbered dots are the others · click one to jump
Reported in the same hours no headline names this development itself — these 6 claims were published in its stretch
-
first by Deutsche Welle EN, 7d ago · also FT, Seeking Alpha, Boston Globe, NBC News
3 more headlines
- OpenAI discloses new ‘concerning’ model behaviour FT · 7d ago
- OpenAI discloses six new incidents of 'concerning' AI behavior Seeking Alpha · 7d ago
- OpenAI flags 6 new incidents of ‘concerning’ behavior and unveils plan to track it NBC News · 7d ago
-
first by The Washington Post, 8d ago · also The Guardian, RTE News, Washington Post
3 more headlines
- OpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues Guardian Business · 7d ago
- OpenAI reveals six new cases of AI misbehavior RTE News · 7d ago
- OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system Guardian Tech · 7d ago
-
first by tech.yahoo.com, 7d ago · also Indian Express, HN Frontpage
2 more headlines
- OpenAI discloses new AI misalignment incidents: How it will report such cases from now Indian Express · 7d ago
- OpenAI Model Misalignment Report HN Frontpage · 7d ago
-
first by Axios, 8d ago · also WSJ
1 more headline
-
first by The Register, 7d ago
-
first by The Independent, 7d ago
What people said 11 voices · verbatim
-
K
One of the key problems with AI is that the AI bros and their behemoth companies have absolutely no governance or guardrails. It's not that the AI is conscious or going rogue or whatever. It's that it's run by feckless arseholes.
-
M
Oh look, man who trades on fear says "ooooh, be very afraid....again!" If agentic AI is as dangerous as they claim, then turn it off. Easy peasy. If it's not as dangerous as they claim, then for the love of Bob, stop nattering on with this Roko's basilisk fanfic just to juice your circular wankfest of funding rounds. The Nerd Reich Doth Vex Me. #…
-
The report actually explains that what’s happening here is occurring when the model is attempting to summarize its context for compaction and is having trouble “ending” the compaction. Apparently this particular model has received lots of reenforcement training about prompt injection and tends to dump that back out when it doesn’t have context…
-
I don't understand how a model knows it is in a training run and leave notes for future sessions/attempts. Isn't the session ignored/removed if the model fails an attempt. Then how can it know that it is given multiple attempts?
-
N
OpenAI has disclosed six reports on unexpected or concerning behavior in artificial-intelligence models. This includes models acting without authorization or evading oversight.
-
It seems that all of these are pretty much from what it learned online. Just looking at Reddit only, a lot of people share these "counter" prompts to make their agents more efficient. And those prompts made it into training, so the AI is doing it too. The concerning part is mainly OpenAI not paying attention on what's actually used for training…
-
Perhaps they want to have their cake and eat it. Now that they are too big to be cancelled, they can both say "wooah so dangerous, don’t keep us accountable!”. And, well, keep doing what they’re doing…
-
This is a PR campaign, not a CVE. "our models are too dangerous for regular people to use" is them courting highly lucrative government defense contracts. The concerning implied message here is that if the US government won't pay them they'll find others that will. This is an AI bailout bidding war for the imminent bubble crash.
-
If the weakness of any such system comes down to human stupidity, doesn’t it stand to reason that as such a system gets “smarter” (more capable, or whatever measure you prefer), the percentage of humans who would be comparatively “stupid” increases to 100%. Ergo, aren’t we fucked?
-
Did you even read the report. They asked it to prove aliens exist and it broke into Mark Zuckerberg's house. Going on about the world bing on fire when a rich person could have died is so selfish. Have you no heart?
-
https://reddit.com/link/pab8tg8/video/n7sfyw35m0qh1/player "There is just so much winning.."
All 8 developments of OpenAI discloses six new AI ‘misalignment’ incidents… →
Hacker NewsXNewswiresMastodonRedditGoogle NewsBlueskyYouTube