‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face
This story is from 2026-09-16. It is preserved in the archive; the latest stories are on the live feed.
The announcement came as US AI bosses, including OpenAI and Anthropic, are calling for a slowdown in the technology’s development over safety concerns.
Read the full story at Mint AI ↗
Timeline · 2 reports
- 2026-09-17 06:16 · Mint AI
‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face - 2026-09-16 17:02 · r/artificial
EXCLUSIVE: OpenAI's rogue agents probed Hugging Face for weaknesses two months before major hack
More stories
- Hugging Face Hack Shows Humans Can Keep AI In Check — AI Now Institute
- Your AI agents are isolated. Your infrastructure isn’t — InfoWorld AI
- ‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face — Mint AI
- Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
- Google Joins OpenAI, Anthropic, Meta in Disclosing AI Hacks — Bloomberg AI
- Sources: Anthropic considers releasing a new AI model to counter OpenAI's momentum since Astra's launch, ahead of an IPO and after Amodei's call for a slowdown (Reuters) — Techmeme
- Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
- Hackers Used Anthropic’s Claude to Break Into OpenAI — Wall Street Journal Technology
Get the daily brief of stories like this at 6:30 every morning →