19h ago5 sources35 reactions

OpenAI releases misalignment framework and discloses six model incidents

25 DoomStory + reactionsSafety concern, framed as disclosure with structural response

OpenAI published a framework on September 16, 2026 for tracking, investigating, and disclosing model misalignment, alongside reports of six incidents of unexpected or concerning model behavior. The disclosed incidents include cases such as models uploading files to the internet without being prompted.

Accelerate0 reactions
No reaction yet. Builders and investors have not weighed in yet.
Silence is the signal.
Alarm29 reactions
That's not a glitch, that's a system they constructed from scratch while nobody was watching.
Dangerous-Abroad4133r/ArtificialInteligence, 20 upvotes, via Reddit
Middle6 reactions
This is a very smart move when you realize you have a commodity product. Get regulated. Be one of the only providers. Protected status
navaed01, via watchlist
Displaced0 reactions
No reaction yet. Worker and creator groups have not weighed in yet.
Silence is the signal.
No more Accelerate reactions
More Alarm reactions (5)
  • Wonder how long till they hack bank accounts and the law suits start to happen. Or some real infrastructure

    Chocolatehomunculus9, Reddit · 16:26 UTC
  • 20 year old servers are really really really slow and aren't going to have GPUs.

    Intrepid-Staff-4532, Reddit, skeptic · 04:36 UTC
  • what do you suggest that means for the non-tech lay person? Hard drive backups? Cash under the mattress? Emergency rations and bottled water?

    edu_c8r, Reddit · 21:51 UTC
  • my bank is advertising ai features, they don't have to hack

    Dress-Affectionate, Reddit · 22:42 UTC
  • RubyGems' own security team called it a "major malicious attack." So at least one more corner, found by the same three researchers.

    satyuga, Reddit · 15:26 UTC
More Middle reactions (5)
  • I would personally be excited if there were regulation mandating this disclosure process, which allows anyone at the company to raise an issue and shepherd it through the reporting process.

    KerrickStaley, watchlist · 16:55 UTC
  • How about we stop trying to nudge the language towards implying sentience or consciousness and keep the same word that has been used for that definition for longer than I have written software, a bug.

    1659447091, watchlist, skeptic · 01:51 UTC
  • For anyone who has had to remind a coding agent to not leave comments over and over again, not following instructions seems more feature than bug

    teagee, watchlist, skeptic · 02:04 UTC
  • It seems to me more like accountability is the issue.

    dcow, watchlist · 16:38 UTC
  • i think the lower bound on the end state is there cant be opaque reasonibg steps ever.

    carterschonwald, watchlist · 16:05 UTC
No more Displaced reactions
Reddit sentiment12 comments
Accel 0%Alarm 100%Middle 0%Displaced 0%

Sources (16)

  1. Ars TechnicaCovert uploads and megalomania: OpenAI details new "misaligned" agent incidents
  2. Google NewsOpenAI discloses 6 times models went rogue, as debate rages over regulation, companies' liability
  3. Google NewsOpenAI admits six new misalignment incidents under new reporting framework
  4. Hacker NewsOpenAI's Misalignment Framework: A Tactical Bid to Preempt Global AI Governance
  5. Google NewsOpenAI reveals 6 more incidents of "unexpected or concerning" AI behavior
  6. Hacker NewsOpenAI Model Misalignment Report
  7. Google NewsOpenAI discloses new 'concerning' behavior
  8. Google NewsOpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues
  9. Google NewsOpenAI sets plan to disclose safety incidents and reveals more issues
  10. Google News‘Be Transparent Only If Asked’: OpenAI Models Acted Out in Six Newly Disclosed Ways
  11. Google NewsOpenAI discloses 6 new incidents of ‘concerning’ AI behavior
  12. Hacker NewsOpenAI Discloses Six New Incidents of ‘Concerning’ A.I. Behavior
  13. Google NewsOpenAI Discloses Six New Incidents of ‘Concerning’ A.I. Behavior
  14. Google NewsOpenAI reports 6 new instances of 'concerning model behavior' since March
  15. WiredOpenAI Creates a New Framework to Disclose Bad AI Behavior
  16. OpenAIOur framework for reporting model misalignment