GTR: a pure linear-attention backbone for six real-time vision tasks
This story is from 2026-09-24. It is preserved in the archive; the latest stories are on the live feed.
GTR 🏎️ is purely recurrent — 12 gated linear attention blocks, scanning in four directions. The video shows five of the six nuScenes tasks we tested. None of these models saw nuScenes during training. We’ve put the whole stack out under MIT: code, weights, CUDA kernel, TensorRT plugin, plus deploy…
Read the full story at r/learnmachinelearning ↗
Timeline · 1 report
- 2026-09-24 14:05 · r/learnmachinelearning
GTR: a pure linear-attention backbone for six real-time vision tasks