RTX 5090 Bonsai 2 27B vs Gemma 4 12B vs Qwen 3.5 9B Japanese voxel pagoda
Yesterday I saw that prismml just dropped the new bonsai which is a heavily-quantized version of qwen 3.8 27b claiming over 98% top-1% comparing to fp16, while weighing from ~6gb at Q1 to ~8gb at Q2 . I wanted to see how good this model is compared to other models in this memory range and my choice…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-18 17:40 · r/LocalLLM
RTX 5090 Bonsai 2 27B vs Gemma 4 12B vs Qwen 3.5 9B Japanese voxel pagoda