Quoting Anthropic Frontier Red Team
We evaluate several models on 100 tasks from the [internal Binary Exploitation benchmark] (selected at random), and find that GLM-5.3 develops full control flow hijacks in 4% of the trials; Claude Mythos Preview did so in 6%. Although GLM-5.3 performs below Claude Mythos Preview here, a meaningful…
Read the full story at Simon Willison's Weblog ↗
Timeline · 1 report
- 2026-09-29 22:20 · Simon Willison's Weblog
Quoting Anthropic Frontier Red Team