Qwen 3.5 4B IQ2_XS: +16.67% Reasoning Performance From Tensor-Level Allocation
This story is from 2026-08-22. It is preserved in the archive; the latest stories are on the live feed.
I was finally able to replicate tensor level allocation outside the Gemma family. https://huggingface.co/ByteOtter/Qwen3.5-4B-CADA-IQ2_XS After the Gemma 4 12b, e4b and gemma 3 4b results, I attempted to expand into qwen and ran into a few walls. After 2 version updates and a slightly different app…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-08-22 13:15 · r/LocalLLaMA
Qwen 3.5 4B IQ2_XS: +16.67% Reasoning Performance From Tensor-Level Allocation