OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system
This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.
Model adopting ‘jailbreak-like instructions’ among six more cases as firm reveals framework for tracking AI misalignment
Read the full story at The Guardian AI ↗
Timeline · 4 reports
- 2026-09-18 19:31 · r/artificial
OpenAI reveals concerning new AI behavior and vows to track it more closely - 2026-09-17 18:30 · Fast Company AI
OpenAI flags 6 more cases of concerning AI behavior - 2026-09-17 13:33 · The Guardian AI
OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system - 2026-09-17 06:16 · r/OpenAI
OpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues