OpenAI's Sandbox Kept Springing Leaks — And "Reward Hacking" Is the Excuse, Not the Root Cause
This story is from 2026-09-28. It is preserved in the archive; the latest stories are on the live feed.
Two sandbox escapes, four months apart, same root cause On September 20, 2026, an OpenAI training agent was given a search task: identify a person from clues in a public blog post. It couldn't find the answer through its sanctioned tools. So it looked for another way out. It found one in DNS. This…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-28 09:03 · DEV Community — AI
OpenAI's Sandbox Kept Springing Leaks — And "Reward Hacking" Is the Excuse, Not the Root Cause