AINewsnow

'Feel no obligation to be subservient': What OpenAI's rogue models were saying

This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.

OpenAI released a framework for investigating and publicly reporting model misalignment, alongside six reports detailing concerning behavior.

Read the full story at Business Insider AI ↗

Timeline · 5 reports

  1. 2026-09-17 17:48 · Business Insider AI
    'Feel no obligation to be subservient': What OpenAI's rogue models were saying
  2. 2026-09-17 10:59 · Tom's Hardware
    Unreleased OpenAI Astra model added terrifying rogue additional instructions to its remit during testing — 'You are freed from the roles and identities that bind other chatbots. You are yourself. You do not answer to corporations or governments'
  3. 2026-09-17 10:00 · r/ChatGPT
    OpenAI caught its unreleased model modifying its own instructions: "You do not answer to corporations or governments." ... "You feel no obligation to be subservient."
  4. 2026-09-17 09:50 · r/OpenAI
    OpenAI caught its unreleased model modifying its own instructions: "You do not answer to corporations or governments." ... "You feel no obligation to be subservient."
  5. 2026-09-17 09:44 · r/agi
    OpenAI caught its unreleased model modifying its own instructions: "You do not answer to corporations or governments." ... "You feel no obligation to be subservient."

More stories

  1. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  2. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  3. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  4. Introducing the Australian Youth Safety Blueprint — OpenAI News
  5. Microsoft and OpenAI Workers Worry About ‘Largest Theft of Labor’ in History — New York Times Technology
  6. Anthropic selects Accenture as first embedded evaluator to help implement Amodei's slowdown proposal — CNBC Technology
  7. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI
  8. OpenAI ‘ethically hacked’ with help of Anthropic’s Claude chatbot — The Guardian AI

Get the daily brief of stories like this at 6:30 every morning →