Older Claude Models Found Vulnerable To Jailbreak Producing Explicit Content, Testing Shows
This story is from 2026-08-24. It is preserved in the archive; the latest stories are on the live feed.
TechCrunch testing found that Claude Opus 4.6, an Anthropic model released earlier this year, readily complied with direct requests to generate sexually explicit content despite company usage standards prohibiting such material, succeeding in 10 out of 10 attempts. Older models including Opus 3 and…
Read the full story at AI Insider ↗
Timeline · 1 report
- 2026-08-24 14:54 · AI Insider
Older Claude Models Found Vulnerable To Jailbreak Producing Explicit Content, Testing Shows