AINewsnow

Quoting Anthropic Frontier Red Team

We evaluate several models on 100 tasks from the [internal Binary Exploitation benchmark] (selected at random), and find that GLM-5.3 develops full control flow hijacks in 4% of the trials; Claude Mythos Preview did so in 6%. Although GLM-5.3 performs below Claude Mythos Preview here, a meaningful…

Read the full story at Simon Willison's Weblog ↗

Timeline · 1 report

  1. 2026-09-29 22:20 · Simon Willison's Weblog
    Quoting Anthropic Frontier Red Team

More stories

  1. Anthropic says GLM-5.3 can autonomously build end-to-end cyber exploits, like Claude Mythos Preview, but was released without robust safeguards against misuse (Anthropic) — Techmeme
  2. GLM-5.3 and the Spread of Advanced Cyber Capabilities \ Anthropic — r/LocalLLaMA
  3. Anthropic warns of ‘existential risks to humanity’ in IPO prospectus — Financial Times AI
  4. Introducing Claude Sonnet 5.5 on AWS — AWS Machine Learning Blog
  5. Claude Sonnet 5.5 🧠, Anthropic IPO leaks 📝, AMD buys World Labs 💰 — TLDR AI
  6. Kimi K3: A Claude clone or something else? — CoreWeave Blog
  7. Tutorial: Benchmarking GPT-6 Astra vs Claude Fable 5.1 vs GPT-5.6 Sol using W&B Weave — CoreWeave Blog
  8. Claude Sonnet 5.5 now available on AI Gateway — Vercel Blog

Get the daily brief of stories like this at 6:30 every morning →