What Attention Recalls and Recurrence Controls in Hybrid Language Models
This story is from 2026-09-07. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.04434v1 Announce Type: new Abstract: Hybrid language models combine attention with a fixed-size recurrent state, but the role of each channel remains unclear. We introduce two cache-level interventions. Split-prefill keeps only the KV cache or only the recurrent state from a prefilled co…
Read the full story at arXiv cs.CL ↗
Timeline · 1 report
- 2026-09-07 04:00 · arXiv cs.CL
What Attention Recalls and Recurrence Controls in Hybrid Language Models