Now you can grow Bonsai on your potato
No more excuses for GPU-poor folks not to start LLMing! Full 27B-class reasoning in ternary transformer weights, for llama.cpp (CUDA, Metal, CPU) ~9.3x smaller than FP16 (ideal) | 98.2% of FP16 intelligence retained | ~47 tok/s on an Apple M5 Max laptop Highlights ~5.9 GB language model (down from…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-11 12:58 · r/LocalLLaMA
Now you can grow Bonsai on your potato