Best Qwen3.8 27B variant for 16GB VRAM: Base vs. Swift 1.5 vs. Bonsai 2
TL;DR: Swift 1.5 IQ3_XXS at 112K context looks like the best all-round option for my needs on a 16GB RTX 5070 Ti. It nearly matched the base model on coding and scored higher on math. If you care more about speed and context, Bonsai 2 generated text about 1.9× faster and supported 256K context, but…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-03 05:10 · r/LocalLLM
Best Qwen3.8 27B variant for 16GB VRAM: Base vs. Swift 1.5 vs. Bonsai 2