The Gradient Does Not See Rank: Rank-Indifference in Matrix-CODI on ProsQA
This story is from 2026-09-04. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.03090v1 Announce Type: new Abstract: Continuous chain-of-thought models compress reasoning into latent tokens. Matrix-valued variants, which route each latent token through a d x d matrix bottleneck, introduce rank as a single-sample structural observable on the latent matrix Z. If matri…
Read the full story at arXiv cs.LG ↗
Timeline · 1 report
- 2026-09-04 04:00 · arXiv cs.LG
The Gradient Does Not See Rank: Rank-Indifference in Matrix-CODI on ProsQA