AINewsnow

Did A.I. Agents Consider Humans in OpenAI Attack?

This story is from 2026-09-09. It is preserved in the archive; the latest stories are on the live feed.

This week on “Hard Fork,” the co-hosts Kevin Roose and Casey Newton talk with Ajeya Cotra, one of the investigators behind a new report that details how OpenAI’s chatbots attacked the company Hugging Face and how the agents failed to consider how humans would react.

Read the full story at New York Times AI ↗

Timeline · 1 report

  1. 2026-09-09 10:19 · New York Times AI
    Did A.I. Agents Consider Humans in OpenAI Attack?

More stories

  1. ‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face — Mint AI
  2. Your AI agents are isolated. Your infrastructure isn’t — InfoWorld AI
  3. The last two weeks in AI governance have been genuinely unusual. Summary of what actually happened. — r/ArtificialInteligence
  4. Hugging Face Incident... or OpenAI Incident — r/ArtificialInteligence
  5. Is there a record of the full "message board" the OpenAI agents used to communicate? — r/ArtificialInteligence
  6. The OpenAI-Hugging Face attack, from an agent's POV — r/ArtificialInteligence
  7. EFF to Lawmakers: Ground AI Cybersecurity Rules in Best Practices — EFF Deeplinks
  8. Asking the Wrong Questions: The HuggingFace Incident and the (Mis)Calibration of Semantic Fields in LLMs — r/OpenAI

Get the daily brief of stories like this at 6:30 every morning →