AINewsnow

GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benchmark

Leading AI models usually attempt dangerous tasks rather than refuse them when controlling a robot, according to the RoboHarm benchmark. GPT-6 Astra stabbed a baby doll in 17 of 20 trials, while Claude Fable 5.1 put a can of compressed air on a burning stove. None of the three models tested reliabl…

Read the full story at The Decoder ↗

Timeline · 1 report

  1. 2026-09-19 13:28 · The Decoder
    GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benchmark

More stories

  1. we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI
  2. I ran Claude code and Codex in parallel for 15 days. Here's what I found. — r/AI_Agents
  3. I built an iOS app with Claude code to break out of my usual chord habits and unlock new progressions. — r/ClaudeAI
  4. Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
  5. This Ford exec put her family's Claude assistant on a PIP. ChatGPT has taken over. — Business Insider AI
  6. AI cybersecurity risks explode as Claude used to break into ChatGPT — Semafor Technology
  7. The cloud outage that should terrify the CIO — InfoWorld AI
  8. 5090, 9850x3d, 64gb ram, where do I get started? — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →