AINewsnow

AI Security Patches Fail 74% of the Time: 6,080 Tested

This story is from 2026-09-11. It is preserved in the archive; the latest stories are on the live feed.

TL;DR 1Password's Off-by-1 Labs had ChatGPT 5.5 and Claude Opus 4.8 write 6,080 patches for six real CVEs. Only 26.0% were complete fixes that did not change how the application behaved. The common failure is not a broken patch. It is a narrow input check that looks like a fix, passes tests, and le…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-11 17:44 · DEV Community — AI
    AI Security Patches Fail 74% of the Time: 6,080 Tested

More stories

  1. Tested Cursor, Claude Code, Codex and Antigravity on the exact same app build — r/AI_Agents
  2. Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
  3. Dumbest solution to the alignment problem — r/singularity
  4. AI cybersecurity risks explode as Claude used to break into ChatGPT — Semafor Technology
  5. The cloud outage that should terrify the CIO — InfoWorld AI
  6. One prompt two models — r/AI_Agents
  7. Solving image to text captchas — r/AI_Agents
  8. Own 1 dashboard for ChatGPT, Gemini, Claude, and more for only $54.97 — Mashable AI

Get the daily brief of stories like this at 6:30 every morning →