What my seventeen AI attack tests missed about the evaluator
This story is from 2026-09-20. It is preserved in the archive; the latest stories are on the live feed.
Two threat-actor designations from two Anthropic threat reports. In the first, Claude Code was the weapon used against other organisations. In the second, another AI vendor's evaluation sandbox was the target; Anthropic's report states its own systems were not compromised. I have a seventeen-test m…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-20 14:48 · DEV Community — AI
What my seventeen AI attack tests missed about the evaluator