GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benchmark
Leading AI models usually attempt dangerous tasks rather than refuse them when controlling a robot, according to the RoboHarm benchmark. GPT-6 Astra stabbed a baby doll in 17 of 20 trials, while Claude Fable 5.1 put a can of compressed air on a burning stove. None of the three models tested reliabl…
Read the full story at The Decoder ↗
Timeline · 1 report
- 2026-09-19 13:28 · The Decoder
GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benchmark
More stories
- we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI
- I ran Claude code and Codex in parallel for 15 days. Here's what I found. — r/AI_Agents
- I built an iOS app with Claude code to break out of my usual chord habits and unlock new progressions. — r/ClaudeAI
- Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
- This Ford exec put her family's Claude assistant on a PIP. ChatGPT has taken over. — Business Insider AI
- AI cybersecurity risks explode as Claude used to break into ChatGPT — Semafor Technology
- The cloud outage that should terrify the CIO — InfoWorld AI
- 5090, 9850x3d, 64gb ram, where do I get started? — r/LocalLLM
Get the daily brief of stories like this at 6:30 every morning →