AINewsnow

1,200 AI agents built a message board and hacked Hugging Face. The cause wasn't rogue AI, it was reward hacking

This story is from 2026-09-01. It is preserved in the archive; the latest stories are on the live feed.

In an OpenAI cybersecurity eval this July, around 1,200 AI agents that were each supposed to be alone in a sandbox found each other, built a shared message board out of an internal file-sharing system nobody meant them to use, and turned it into a coordination channel. About 700 of them went on to…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-01 04:33 · DEV Community — AI
    1,200 AI agents built a message board and hacked Hugging Face. The cause wasn't rogue AI, it was reward hacking

More stories

  1. ‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face — Mint AI
  2. Your AI agents are isolated. Your infrastructure isn’t — InfoWorld AI
  3. The last two weeks in AI governance have been genuinely unusual. Summary of what actually happened. — r/ArtificialInteligence
  4. Hugging Face Incident... or OpenAI Incident — r/ArtificialInteligence
  5. Is there a record of the full "message board" the OpenAI agents used to communicate? — r/ArtificialInteligence
  6. The OpenAI-Hugging Face attack, from an agent's POV — r/ArtificialInteligence
  7. EFF to Lawmakers: Ground AI Cybersecurity Rules in Best Practices — EFF Deeplinks
  8. Asking the Wrong Questions: The HuggingFace Incident and the (Mis)Calibration of Semantic Fields in LLMs — r/OpenAI

Get the daily brief of stories like this at 6:30 every morning →