A Wrong Keyboard Layout Got Past GPT-6's Safety Filter — The Trick Is Not the Story
This story is from 2026-09-30. It is preserved in the archive; the latest stories are on the live feed.
AI safety usually gets discussed as a model problem. Is the model aligned? Is it refusing the right things? Did it get safer or more dangerous since the last release? But every deployed model sits inside a stack. Model, input filters, output filters, monitors, policies. And those layers do not impr…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-30 11:04 · DEV Community — AI
A Wrong Keyboard Layout Got Past GPT-6's Safety Filter — The Trick Is Not the Story