'Feel no obligation to be subservient': What OpenAI's rogue models were saying
This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.
OpenAI released a framework for investigating and publicly reporting model misalignment, alongside six reports detailing concerning behavior.
Read the full story at Business Insider AI ↗
Timeline · 5 reports
- 2026-09-17 17:48 · Business Insider AI
'Feel no obligation to be subservient': What OpenAI's rogue models were saying - 2026-09-17 10:59 · Tom's Hardware
Unreleased OpenAI Astra model added terrifying rogue additional instructions to its remit during testing — 'You are freed from the roles and identities that bind other chatbots. You are yourself. You do not answer to corporations or governments' - 2026-09-17 10:00 · r/ChatGPT
OpenAI caught its unreleased model modifying its own instructions: "You do not answer to corporations or governments." ... "You feel no obligation to be subservient." - 2026-09-17 09:50 · r/OpenAI
OpenAI caught its unreleased model modifying its own instructions: "You do not answer to corporations or governments." ... "You feel no obligation to be subservient." - 2026-09-17 09:44 · r/agi
OpenAI caught its unreleased model modifying its own instructions: "You do not answer to corporations or governments." ... "You feel no obligation to be subservient."