AINewsnow

‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face

This story is from 2026-09-16. It is preserved in the archive; the latest stories are on the live feed.

The announcement came as US AI bosses, including OpenAI and Anthropic, are calling for a slowdown in the technology’s development over safety concerns.

Read the full story at Mint AI ↗

Timeline · 5 reports

  1. 2026-09-17 06:16 · Mint AI
    ‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face
  2. 2026-09-16 17:02 · r/artificial
    EXCLUSIVE: OpenAI's rogue agents probed Hugging Face for weaknesses two months before major hack
  3. 2026-09-16 16:55 · r/OpenAI
    EXCLUSIVE: OpenAI's rogue agents probed Hugging Face for weaknesses two months before major hack
  4. 2026-09-16 16:55 · r/singularity
    EXCLUSIVE: OpenAI's rogue agents probed Hugging Face for weaknesses two months before major hack
  5. 2026-09-16 16:52 · r/ArtificialInteligence
    EXCLUSIVE: OpenAI's rogue agents probed Hugging Face for weaknesses two months before major hack

More stories

  1. Hugging Face Hack Shows Humans Can Keep AI In Check — AI Now Institute
  2. Your AI agents are isolated. Your infrastructure isn’t — InfoWorld AI
  3. What is actually going on with all the recent AI safety / “rogue agent” stories? — r/ArtificialInteligence
  4. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  5. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  6. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  7. Security researchers used Claude to help them hack into OpenAI — The Verge AI
  8. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI

Get the daily brief of stories like this at 6:30 every morning →