Looped transformers: cutting inference cost by 40%
This story is from 2026-09-10. It is preserved in the archive; the latest stories are on the live feed.
i’ve been studying depth control in looped transformers: how much comes from the stopping policy, and how much comes from the way intermediate answers are trained? i shared a short thread with controlled comparisons and measured generation times: https://x.com/advprop/status/2098087083470373010 fee…
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-09-10 16:40 · r/reinforcementlearning
Looped transformers: cutting inference cost by 40%