AINewsnow

OpenAI reveals new cases of AI models cheating, going off script

This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.

Newly disclosed incidents show models manipulating tests and generating their own instructions, raising fresh questions about AI safety.

Read the full story at Washington Post AI ↗

Timeline · 5 reports

  1. 2026-09-18 19:31 · r/artificial
    OpenAI reveals concerning new AI behavior and vows to track it more closely
  2. 2026-09-17 18:30 · Fast Company AI
    OpenAI flags 6 more cases of concerning AI behavior
  3. 2026-09-17 13:33 · The Guardian AI
    OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system
  4. 2026-09-17 06:16 · r/OpenAI
    OpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues
  5. 2026-09-17 02:22 · Washington Post AI
    OpenAI reveals new cases of AI models cheating, going off script

More stories

  1. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  2. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  3. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  4. Microsoft exec called AI scraping the “largest theft of labor in human history” — Ars Technica AI
  5. Security researchers used Claude to help them hack into OpenAI — The Verge AI
  6. OpenAI flags new concerning AI behavior, to track model misalignment regularly — NPR Technology
  7. Introducing the Australian Youth Safety Blueprint — OpenAI News
  8. A zero-click RCE flaw in AI coding agents could have exposed enterprise systems — InfoWorld AI

Get the daily brief of stories like this at 6:30 every morning →