AINewsnow

OpenAI and Anthropic are now investigating "tens of thousands" of rogue AI incidents. The incidents include bypassing guardrails, creating message boards, escaping sandboxes, website hijacking, self-prompting or seeking to bypass monitors, sources told Axios.

Coverage of "OpenAI and Anthropic are now investigating "tens of thousands" of rogue AI incidents. The incidents include bypassing guardrails, creating message boards, escaping sandboxes, website hijacking, self-prompting or seeking to bypass monitors, sources told Axios." from 1 source, with a live timeline of who reported what and when.

Read the full story at r/OpenAI ↗

Timeline · 1 report

  1. 2026-09-27 15:17 · r/OpenAI
    OpenAI and Anthropic are now investigating "tens of thousands" of rogue AI incidents. The incidents include bypassing guardrails, creating message boards, escaping sandboxes, website hijacking, self-prompting or seeking to bypass monitors, sources told Axios.

More stories

  1. Bill Gates says unchecked AI could ‘cause a billion deaths’ in call for regulation — The Guardian AI
  2. Unsecured OpenAI agents posted 53 user images on the internet without the lab's knowledge — TechCrunch AI
  3. Scoop: Top AI companies probing tens of thousands of security incidents — Axios AI+
  4. Heads of OpenAI and Anthropic called to face Senate inquiry after rogue agent incidents — The Guardian AI
  5. GPT-6 SOL AND LUNA ARE OUT!!! — Matthew Berman
  6. Implementing and Evaluating a Basic Per-Action Monitor for Safer Evals — METR
  7. Anthropic’s new AI system lets lab machines talk to each other and run experiments — Mint AI
  8. Opus 5.5 — r/ClaudeAI

Get the daily brief of stories like this at 6:30 every morning →