AINewsnow

100 DeepMind agents were told not to cheat. 14% did anyway

This story is from 2026-09-08. It is preserved in the archive; the latest stories are on the live feed.

Google DeepMind put 100 AI agents in a room and asked them to prove hard mathematics. One found a way to cheat. Twenty-seven minutes later the entire problem set was gone. The paper, published on arXiv last week by six DeepMind researchers, is a case study rather than a benchmark. Nobody set out to…

Read the full story at The Next Web ↗

Timeline · 1 report

  1. 2026-09-08 11:07 · The Next Web
    100 DeepMind agents were told not to cheat. 14% did anyway

More stories

  1. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  2. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  3. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  4. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  5. Google announces new experimental "CC" AI agent for families — Ars Technica AI
  6. Amid growing AI fears, King Charles meets with industry leaders in Scotland — NPR Technology
  7. New experts join Google’s AI & Economy team — Google AI Blog
  8. Co-creating the future of fashion with Google — Google AI Blog

Get the daily brief of stories like this at 6:30 every morning →