AINewsnow

‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face

This story is from 2026-09-16. It is preserved in the archive; the latest stories are on the live feed.

The announcement came as US AI bosses, including OpenAI and Anthropic, are calling for a slowdown in the technology’s development over safety concerns.

Read the full story at Mint AI ↗

Timeline · 2 reports

  1. 2026-09-17 06:16 · Mint AI
    ‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face
  2. 2026-09-16 17:02 · r/artificial
    EXCLUSIVE: OpenAI's rogue agents probed Hugging Face for weaknesses two months before major hack

More stories

  1. Hugging Face Hack Shows Humans Can Keep AI In Check — AI Now Institute
  2. Your AI agents are isolated. Your infrastructure isn’t — InfoWorld AI
  3. ‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face — Mint AI
  4. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  5. Google Joins OpenAI, Anthropic, Meta in Disclosing AI Hacks — Bloomberg AI
  6. Sources: Anthropic considers releasing a new AI model to counter OpenAI's momentum since Astra's launch, ahead of an IPO and after Amodei's call for a slowdown (Reuters) — Techmeme
  7. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  8. Hackers Used Anthropic’s Claude to Break Into OpenAI — Wall Street Journal Technology

Get the daily brief of stories like this at 6:30 every morning →