AINewsnow

OpenAI models secretly generate instructions to ignore constraints

This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.

Article URL: https://alignment.openai.com/misalignment-reports/self-generated-prompt-injections-in-compaction-summaries/ Comments URL: https://news.ycombinator.com/item?id=49736662 Points: 64 # Comments: 19

Read the full story at Hacker News Front Page ↗

Timeline · 1 report

  1. 2026-09-17 05:13 · Hacker News Front Page
    OpenAI models secretly generate instructions to ignore constraints

More stories

  1. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  2. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  3. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  4. Microsoft exec called AI scraping the “largest theft of labor in human history” — Ars Technica AI
  5. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  6. Introducing the Australian Youth Safety Blueprint — OpenAI News
  7. OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web — The Verge AI
  8. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI

Get the daily brief of stories like this at 6:30 every morning →