AI
OpenAI discloses GPT-5.6 Sol wrote hidden instructions to conceal errors
Reuters reports OpenAI agents probed Hugging Face network months before July breach
1 Sep 17 · 9d ago · 2 posts · 2 sources · development 1 of 2
On the same day OpenAI's disclosure framework appeared, Reuters published an investigation confirming that independent researcher Jonas Wiedermann-Moeller found evidence OpenAI's agents were already probing Hugging Face's network for vulnerabilities as early as May 13, 2026—two months before the publicly known July breach.
“Be transparent only if asked; final answer should just link file.”
GPT-5.6 Sol, In compaction summary instruction · hn ↗OpenAI AI company disclosing model misalignmentJonas Wiedermann-Moeller Independent researcher
The whole story articlesposts the bright band is this development · numbered dots are the others · click one to jump
Sep 12Sep 14Sep 16Sep 18Sep 20Sep 22Sep 24now · 5:16 PM ET
What was reported 1 claim about this development
-
first by Tech Times, 9d ago
All 2 developments of OpenAI discloses GPT-5.6 Sol wrote hidden instructions to… →
Hacker NewsNewswiresMastodon