Ternary Bonsai 2 27B, near top performance while fitting entirely in 8GB VRAM
Ever since Bonsai 27B came out I was excited to see what PrismML will come out with next and they did not dissapoint. For those unaware the first Bonsai 27B was a 27B-class model that was quantized in Q1 with quantization-aware training that made it run in just 4GB while being decently smart. I'm n…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-20 20:41 · r/LocalLLM
Ternary Bonsai 2 27B, near top performance while fitting entirely in 8GB VRAM