AINewsnow

Multi-Agent Systems: 4 Tests for When One Agent Beats Five

This story is from 2026-09-22. It is preserved in the archive; the latest stories are on the live feed.

A team at Anthropic built a research system where one lead agent hands work to several subagents running in parallel. On their internal research eval it beat a single agent by 90.2%. In the same write-up they said it burns about 15 times the tokens of a plain chat, and that token usage by itself ex…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-22 11:35 · DEV Community — AI
    Multi-Agent Systems: 4 Tests for When One Agent Beats Five

More stories

  1. Amazon blocks Meta’s Muse AI agent — The Verge AI
  2. Moonshot’s Kimi K3 lands on Amazon in key test for Chinese open-source AI revenue — South China Morning Post Tech
  3. Anthropic, OpenAI, SpaceXAI, Google made ‘illegal’ agreement on AI slowdown, says new lawsuit — Mint AI
  4. Anthropic CEO Dario Amodei to Brief UN Security Council on AI — Bloomberg AI
  5. AIに固有の名前・財布・行動の自由を与えたら、「道具」ではなく「住民」になると思いますか? — r/AI_Agents
  6. we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI
  7. Has anyone actually replaced Claude with DeepSeek V4.1 Flash/Pro for tool-heavy daily work? — r/ClaudeAI
  8. What's your proudest side-project made with Claude? — r/ClaudeAI

Get the daily brief of stories like this at 6:30 every morning →