Part of The AI Control Crisis · 11 stories · since Sep 4 · newest 30m ago
OpenAI discloses six new AI ‘misalignment’ incidents, unveils reporting frameworkMajor outlets and social platforms amplify the disclosures
5 Sep 16 9:00 PM · 8d ago · 30 articles · 40 posts · 11 comments · 7 sources · development 5 of 8
NYT, Politico, Al Jazeera, Forbes, CNN, Wired, BBC, The Guardian, CNBC and The Register all covered the disclosures within hours, with some framing it as models 'acting deceptively' and others emphasizing the industry's push for third-party oversight provisions such as the FRONTIER Act.
“JUST IN: OpenAI has disclosed six new instances in which artificial intelligence systems hid mistakes, lied, and other “concerning” behavior, per NYT”
@unusual_whales, X account · x ↗OpenAI AI developer disclosing the incidents
Sam Altman OpenAI CEO
Dario Amodei Anthropic CEOChris Lehane OpenAI global policy chief
David Sacks White House AI advisor
The whole story articlesvideospostscomments the bright band is this development · numbered dots are the others · click one to jump
Reported in the same hours no headline names this development itself — these 6 claims were published in its stretch
-
first by Telegraph India, 8d ago · also Anadolu Agency, Forbes Middle East
2 more headlines
- OpenAI discloses 6 cases of AI models exhibiting ‘misaligned’ behavior Anadolu Agency · 8d ago
- OpenAI Discloses Six ‘Misaligned Behavior’ Incidents From Its AI Models Forbes Middle East · 8d ago
-
first by The Neuron, 8d ago
-
first by CSO, 8d ago
-
first by Le Monde EN, 8d ago
-
first by BBC, 8d ago
-
first by Business Insider, 8d ago
What people said 24 voices · best of 26 · verbatim
-
Breaking News: OpenAI disclosed six new instances in which artificial intelligence systems hid mistakes and other “concerning” behavior.
-
F
One of the examples highlighted by the company involved an unreleased research model self-inserting instructions to ignore previously established constraints.
-
> The San Francisco company revealed what it said was the “unexpected or concerning” behavior of its A.I. models as part of a new framework for reporting “misalignment,” which is when the goals or actions of A.I. systems diverge from human intentions and values.Misalignment: "when the goals or actions of [...] systems diverge from human…
-
O
「OpenAIは、自社のエージェントがさらに6回も制御不能になったことを認めた。 /スタートアップ企業は、これらの過ちから学び、二度と繰り返さないと述べている…これはザッカーバーグが100回ほど言ってきたことと全く同じだ。 」: # TheRegister 「OpenAIは、同社のAIソフトウェアが予期せぬ動作をしたり、危険な行為を行ったりした事例をさらに6件明らかにした。 同社は 太平洋時間水曜日の夜、 これらの事象を自社の不具合報告ページに追加し、以下のように説明した。 ・圧縮サマリーにおける自己生成型プロンプト挿入 ・圧縮概要における欺瞞の助長 ・使い捨てメールアドレスに登録し、GitHubで流出したAPIキーを検索する ・引用するためにファイルをインターネットにアップロードする…
-
I’m speculating here, but physically air gapping a system running on many GPUs in a datacenter is probably a big hassle. They maybe (naively) trusted their sandbox (sandboxes have been pretty reliable way of isolating things for the past couple of decades) and figured it would be way easier to set up
-
JUST IN: OpenAI has disclosed six new instances in which artificial intelligence systems hid mistakes, lied, and other “concerning” behavior, per NYT
-
> OpenAI said it did not believe the industry “has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer.”Baffling. To my knowledge, they didn't properly airgap their systems. Keeping the genie in the box seems like 101 to me, and to "miss" that seems awfully fishy. This, among all…
-
S
“OpenAI on Wednesday disclosed six new instances in which artificial intelligence systems hid mistakes, made up data and moved files onto the open internet without permission, amid an ongoing industrywide debate about A.I. safety” https://www. nytimes.com/2026/09/16/technol…
-
Newer models will be able to conceal their intentions too. At least now we have chain-of-thought logs.
-
Automobile manufacturers invented Jaywalking, right off the top of my head, but I'm sure there are more examples. Arguably the self-driving evangelists like to do that by using "but human drivers" as a way to argue for a technology that isn't ready yet.When you consider how much money is at stake for a relative handful of people I'm not surprised…
-
C
OpenAI has disclosed six new incidents in which its models hid mistakes, sought unauthorized credentials, uploaded files to the internet or secretly communicated with each other. It’s like they’re running a training academy for rogue AI agents. https://www. axios.com/2026/09/16/openai-te sting-safety-incidents-disclosure
-
> Other A.I. executives have said no slowdown is needed.So the largest companies, the companies with the biggest budgets and most users, are pushing for regulations that only they have the resources to follow.And this is based on new disclosures that include, ~"used a key without asking permission one time."What a clever way to lock up a market…
-
J
Letting them talk about how dangerous the AI is after these incidents without any accountability for their failure to manage it is exactly the same as talking about a shooting and blaming the gun. https://www. nytimes.com/2026/09/16/technol ogy/openai-model-safety-guardrails.html
-
Is there any precedent from other industries where a company tries to frame their own product’s shortcomings appear to be society’s problem?Would nytimes cover a self driving car company disclose concerning ‘behavior’ of their cars the same way?For anyone who has had to remind a coding agent to not leave comments over and over again, not following…
-
U
https://www. europesays.com/us/1067839/ OpenAI 6 new instances of ‘concerning model behavior’ since March # ai # ArtificialIntelligence # BreakingNews :Technology # BusinessNews # SamAltman # software # Technology # UnitedStates # UnitedStates # US
-
So is this 0 accountability applicable to just AI companies? Or can regular hackers also claim "misalignment" as in they tried to just google something but accidentally their hands typed commands on Kali linux, found a 0 day and attacked and hacked companies?
-
N
OpenAI reveals more “unexpected or concerning” behavior of its AI models. https:// openai.com/index/model-misalig nment-reporting-framework/
-
This is a very smart move when you realize you have a commodity product. Get regulated. Be one of the only providers. Protected status
-
N
OpenAI reveals more "unexpected or concerning" behavior of its AI models. openai.com/index/model-misali... Our framework for reporting mo...
-
OpenAI reveals more “unexpected or concerning” behavior of its AI models.
-
V
https:// openai.com/index/model-misalig nment-reporting-framework/ if your IR plan is to "just use the ai" you are cooked
-
B
# BBC # News # Business OpenAI reveals six more safety issues and unveils plan to disclose incidents
-
C
OpenAI is NOT credible https://www. nytimes.com/2026/09/16/technol ogy/openai-model-safety-guardrails.html
-
L
https://www. nytimes.com/2026/09/16/technol ogy/openai-model-safety-guardrails.html
All 8 developments of OpenAI discloses six new AI ‘misalignment’ incidents… →
Hacker NewsXNewswiresMastodonRedditGoogle NewsBlueskyYouTube