Looped reasoning means the AI's visible trace isn't the reasoning
This story is from 2026-09-11. It is preserved in the archive; the latest stories are on the live feed.
GPT-6 Astra ships looped transformers: the same blocks run ~44 passes, reusing weights, so effective depth doubles without new parameters. The KV cache and the intermediate states differ per pass, but the tokens you actually see are only one layer of that surface. Here's the part that matters for a…
Read the full story at DEV Community — Machine Learning ↗
Timeline · 1 report
- 2026-09-11 00:15 · DEV Community — Machine Learning
Looped reasoning means the AI's visible trace isn't the reasoning