‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face
This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.
The announcement came as US AI bosses, including OpenAI and Anthropic, are calling for a slowdown in the technology’s development over safety concerns.
Read the full story at Mint AI ↗
Timeline · 1 report
- 2026-09-17 06:16 · Mint AI
‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face
More stories
- Hugging Face Hack Shows Humans Can Keep AI In Check — AI Now Institute
- Your AI agents are isolated. Your infrastructure isn’t — InfoWorld AI
- What is actually going on with all the recent AI safety / “rogue agent” stories? — r/ArtificialInteligence
- Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
- Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
- we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI
- Anthropic selects Accenture as first embedded evaluator to help implement Amodei's slowdown proposal — CNBC Technology
- Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI
Get the daily brief of stories like this at 6:30 every morning →