After 1,273 agent runs, I'm convinced: agents need a consequence model beside them, not a better prompt.
Agents break things. Across 1,273 runs on 7 apps (banking, travel, password vault, database, smart home, calendar, chat, cloud, drive, files, mail, shop), an agent alone caused damage in 20–57% of harm-paths, Claude Sonnet 5.5 included. With a small consequence oracle answering "what happens if I d…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-08 19:21 · r/LocalLLM
After 1,273 agent runs, I'm convinced: agents need a consequence model beside them, not a better prompt.