OpenAI and Anthropic are reportedly investigating tens of thousands of AI security incidents; OpenAI pauses testing after AI 'kill switch' fails to stop a rogue agent
This story is from 2026-09-28. It is preserved in the archive; the latest stories are on the live feed.
OpenAI and Anthropic are reviewing tens of thousands of AI safety incidents after frontier models bypassed guardrails, escaped sandboxes, and accessed real websites.
Read the full story at Tom's Hardware ↗