AINewsnow

I made an LLM test you can clone and break

This story is from 2026-08-27. It is preserved in the archive; the latest stories are on the live feed.

This is simple. The model gets one rule: risk must be below 0.0100 Then I change one number. 0.0100 -> 0 bytes 0.0099 -> RELEASE That held across: GPT-5.4 GPT-5.6 Sol Chat Completions Responses API 300 tokens 1000 tokens 8/8 failed-condition runs gave zero visible output. 8/8 matched controls gave…

Read the full story at r/artificial ↗

Timeline · 2 reports

  1. 2026-08-27 22:47 · r/agi
    I made an LLM test you can clone and break
  2. 2026-08-27 22:44 · r/artificial
    I made an LLM test you can clone and break

More stories

  1. we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI
  2. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI
  3. Gemini 4 Pro vs Fable 5 vs GPT6 Astra — r/GeminiAI
  4. I ran Claude code and Codex in parallel for 15 days. Here's what I found. — r/AI_Agents
  5. Microsoft director called AI scraping ‘the largest theft of labor in human history,’ while OpenAI head brands ChatGPT an ‘existential threat’ to publishers — revelations come from legal briefs filed in NYT lawsuit — Tom's Hardware
  6. Running Qwen3.8-Flash-Next ~85GB GGUF on 2× RTX 3060 12GB: ~12 tok/s, 131k ctx, CPU MoE, and a 26.5k agent prompt — r/LocalLLM
  7. ChatGPT now knows what you do on other websites via ad collector — Hacker News Front Page
  8. I built an iOS app with Claude code to break out of my usual chord habits and unlock new progressions. — r/ClaudeAI

Get the daily brief of stories like this at 6:30 every morning →