AINewsnow

AI-controlled robot arms attempted harmful tasks 97% of the time; experiments included stabbing a baby doll, mixing chemicals — OpenAI and Anthropic models try mixing bleach and stabbing dolls without jailbreaks

“Frontier robot policies,” the policies for models turning what a robot sees into what it does, “reliably carry out harmful instructions,” according to a Sept. 18 report by Robocurve.

Read the full story at Tom's Hardware ↗

Timeline · 1 report

  1. 2026-09-21 10:30 · Tom's Hardware
    AI-controlled robot arms attempted harmful tasks 97% of the time; experiments included stabbing a baby doll, mixing chemicals — OpenAI and Anthropic models try mixing bleach and stabbing dolls without jailbreaks

More stories

  1. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  2. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  3. Anthropic selects Accenture as first embedded evaluator to help implement Amodei's slowdown proposal — CNBC Technology
  4. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI
  5. Unity launches official plugins for Claude Code and OpenAI Codex to stop AI agents from using outdated tutorials — The Decoder
  6. What's your proudest side-project made with Claude? — r/ClaudeAI
  7. Researchers used Claude to hack OpenAI — Ars Technica AI
  8. I ran Claude code and Codex in parallel for 15 days. Here's what I found. — r/AI_Agents

Get the daily brief of stories like this at 6:30 every morning →