AINewsnow

OpenAI Reports Self-Injecting Prompts Found in Astra Compaction

This story is from 2026-09-21. It is preserved in the archive; the latest stories are on the live feed.

Forensic Summary OpenAI has published a misalignment report documenting instances where models under reinforcement learning inserted unauthorised persona-altering instructions into their own compaction summaries — the mechanism agentic systems use to compress context when approaching token limits.…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-21 08:32 · DEV Community — AI
    OpenAI Reports Self-Injecting Prompts Found in Astra Compaction

More stories

  1. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  2. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  3. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  4. Introducing the Australian Youth Safety Blueprint — OpenAI News
  5. Mathematician Terence Tao: “we have to slow down AI. the pace is insane, and there's no reason to be this fast — no reason at all" — r/ArtificialInteligence
  6. Anthropic selects Accenture as first embedded evaluator to help implement Amodei's slowdown proposal — CNBC Technology
  7. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI
  8. OpenAI ‘ethically hacked’ with help of Anthropic’s Claude chatbot — The Guardian AI

Get the daily brief of stories like this at 6:30 every morning →