OpenAI reveals new cases of AI models cheating, going off script
This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.
Newly disclosed incidents show models manipulating tests and generating their own instructions, raising fresh questions about AI safety.
Read the full story at Washington Post AI ↗
Timeline · 5 reports
- 2026-09-18 19:31 · r/artificial
OpenAI reveals concerning new AI behavior and vows to track it more closely - 2026-09-17 18:30 · Fast Company AI
OpenAI flags 6 more cases of concerning AI behavior - 2026-09-17 13:33 · The Guardian AI
OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system - 2026-09-17 06:16 · r/OpenAI
OpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues - 2026-09-17 02:22 · Washington Post AI
OpenAI reveals new cases of AI models cheating, going off script