[2506.13771] LittleBit: Ultra Low-Bit Quantization via Latent Factorization
Interesting to see improvements and research into quantization aware training (QAT) that can make some really tiny models.
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-08 14:23 · r/LocalLLaMA
[2506.13771] LittleBit: Ultra Low-Bit Quantization via Latent Factorization