AINewsnow

We tested if frontier AI robots would refuse dangerous tasks. GPT stabbed a baby doll 17 times, Claude tried to make mustard gas, and the dumbest model was the safest because it kept dropping things.

To see if frontier robot arms have any survival instincts whatsoever, my team and I tested whether the latest AI models would refuse completely unhinged chores like stabbing a baby doll, tossing aerosol cans onto lit stoves, and mixing bleach with ammonia. As it turns out, our future autonomous hel…

Read the full story at r/antiai ↗

Timeline · 1 report

  1. 2026-09-19 09:18 · r/antiai
    We tested if frontier AI robots would refuse dangerous tasks. GPT stabbed a baby doll 17 times, Claude tried to make mustard gas, and the dumbest model was the safest because it kept dropping things.

More stories

  1. AI's role in building AI surging? Anthropic says Claude now leads 26% of its R&D — Mint AI
  2. Claude, Anthropic’s AI model, is helping to develop the next version of itself — Fast Company AI
  3. AI skills — r/AI_Agents
  4. we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI
  5. Anthropic selects Accenture as first embedded evaluator to help implement Amodei's slowdown proposal — CNBC Technology
  6. Bolt Adds DeepSeek V4.1 Flash at 10x Cheaper Than V4 Pro — AlphaSignal
  7. OpenAI ‘ethically hacked’ with help of Anthropic’s Claude chatbot — The Guardian AI
  8. How To Use Ai and create those videos — r/aivideo

Get the daily brief of stories like this at 6:30 every morning →