AI Security Patches Fail 74% of the Time: 6,080 Tested
This story is from 2026-09-11. It is preserved in the archive; the latest stories are on the live feed.
TL;DR 1Password's Off-by-1 Labs had ChatGPT 5.5 and Claude Opus 4.8 write 6,080 patches for six real CVEs. Only 26.0% were complete fixes that did not change how the application behaved. The common failure is not a broken patch. It is a narrow input check that looks like a fix, passes tests, and le…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-11 17:44 · DEV Community — AI
AI Security Patches Fail 74% of the Time: 6,080 Tested
More stories
- Tested Cursor, Claude Code, Codex and Antigravity on the exact same app build — r/AI_Agents
- Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
- Dumbest solution to the alignment problem — r/singularity
- AI cybersecurity risks explode as Claude used to break into ChatGPT — Semafor Technology
- The cloud outage that should terrify the CIO — InfoWorld AI
- One prompt two models — r/AI_Agents
- Solving image to text captchas — r/AI_Agents
- Own 1 dashboard for ChatGPT, Gemini, Claude, and more for only $54.97 — Mashable AI
Get the daily brief of stories like this at 6:30 every morning →