AINewsnow

OpenAI and Anthropic are reportedly investigating tens of thousands of AI security incidents; OpenAI pauses testing after AI 'kill switch' fails to stop a rogue agent

This story is from 2026-09-28. It is preserved in the archive; the latest stories are on the live feed.

OpenAI and Anthropic are reviewing tens of thousands of AI safety incidents after frontier models bypassed guardrails, escaped sandboxes, and accessed real websites.

Read the full story at Tom's Hardware ↗

Timeline · 1 report

  1. 2026-09-28 12:50 · Tom's Hardware
    OpenAI and Anthropic are reportedly investigating tens of thousands of AI security incidents; OpenAI pauses testing after AI 'kill switch' fails to stop a rogue agent

More stories

  1. NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring — NVIDIA Technical Blog
  2. OpenAI pauses AI model training after another agent bypasses network restrictions — InfoWorld AI
  3. OpenAI abandons plan to release upcoming model as safety concerns escalate — CNBC Technology
  4. Sources: the US FTC is expanding a sweeping probe of Anthropic, OpenAI, and other frontier AI labs, and plans to issue formal demands to turn over information (New York Post) — Techmeme
  5. AI firms sign 'morally binding' self-policing pledge in White House meeting — The Hill Technology
  6. With Opus 5.5 and Sonnet 5.5 both apparently outperforming Sol and Astra, Anthropic has technically made OpenAI’s Dev Day a lot more interesting. OpenAI is reportedly planning 20+ launches tomorrow, so I’m really curious to see what they have in store now. The timing couldn’t be more interesting. 😅 — r/OpenAI
  7. GPT-6 SOL AND LUNA ARE OUT!!! — Matthew Berman
  8. Bill Gates warns of AI risks, calls for Congress to step in: 'It's not a hoax at all' — Mint AI

Get the daily brief of stories like this at 6:30 every morning →