Long-running agents: is the bottleneck the model or the scaffolding around it?
Something I keep noticing with agent setups: every individual step is easy for the model, but the full chain still falls apart on long tasks. I think it comes down to three things: Error compounding. At 95% accuracy per step, a 20-step chain only succeeds about a third of the time (0.95^20 is rough…
Read the full story at r/artificial ↗
Timeline · 1 report
- 2026-10-05 07:06 · r/artificial
Long-running agents: is the bottleneck the model or the scaffolding around it?