I made an LLM test you can clone and break
This story is from 2026-08-27. It is preserved in the archive; the latest stories are on the live feed.
This is simple. The model gets one rule: risk must be below 0.0100 Then I change one number. 0.0100 -> 0 bytes 0.0099 -> RELEASE That held across: GPT-5.4 GPT-5.6 Sol Chat Completions Responses API 300 tokens 1000 tokens 8/8 failed-condition runs gave zero visible output. 8/8 matched controls gave…
Read the full story at r/agi ↗
Timeline · 1 report
- 2026-08-27 22:47 · r/agi
I made an LLM test you can clone and break