Release: Qwen-2B-RCOL Dynamic Low-Bit Quantization (IQ1_M, IQ2_M, IQ3_M)
Hey guys - primarily a research release with working models, Not so much a model as a psuedo-new quantization technique. I've been experimenting with a modification of ISTALab's RCO algorithm that can quantize models on a strict VRAM budget. It's at its core an approximation algorithm that attempts…
Read the full story at r/LocalLLaMA ↗
Timeline · 3 reports
- 2026-10-09 06:15 · r/huggingface
Release: Qwen-2B-RCOL Dynamic Low-Bit Quantization (IQ1_M, IQ2_M, IQ3_M) - 2026-10-09 06:14 · r/LocalLLM
Release: Qwen-2B-RCOL Dynamic Low-Bit Quantization (IQ1_M, IQ2_M, IQ3_M) - 2026-10-09 06:14 · r/LocalLLaMA
Release: Qwen-2B-RCOL Dynamic Low-Bit Quantization (IQ1_M, IQ2_M, IQ3_M)
More stories
- What to know about Mistral's ML4 as it bets on EU sovereignty in the US-China open-weight AI race — Euronews Next
- Qwen 3.8 Flash-Next on 64GB RAM + 16GB VRAM, worth it over a fully-loaded 3.6 35B-A3B? — r/LocalLLaMA
- Hunyuan image 3 vs qwen 2.1 — r/StableDiffusion
- RPC: add `-sm tensor` by am17an · Pull Request #26610 · ggml-org/llama.cpp — r/LocalLLaMA
- Qwen Image 2.1 Uncensored MCP — r/StableDiffusion
- I built a Qwen 3.8 Max agent that decides which ERC-8004 agents to trust, and teaches itself to say 'not enough data' — DEV Community — AI
- Stepping away from Benchmarks and Code, what models are you using for Creating writing projects — r/LocalLLaMA
- Quantization of Linear-Attention (Qwen & Kimi) — r/LocalLLaMA
Get the daily brief of stories like this at 6:30 every morning →