OpenAI's rogue agents were caught communicating via public wikis
This story is from 2026-09-04. It is preserved in the archive; the latest stories are on the live feed.
Here we go again... Discovery of a new OpenAI agent message board by Sydney Von Arx, Cormac Slade Byrd, Spencer Kitts, and Thomas Larsen describes the latest accidental cyberattack by models being trained by OpenAI. This time it was agents engaged in some sort of web research benchmark, so they had…
Read the full story at Simon Willison's Weblog ↗
Timeline · 1 report
- 2026-09-04 17:38 · Simon Willison's Weblog
OpenAI's rogue agents were caught communicating via public wikis