A better coder for the small-GPU/small-RAM crowd!
I’ve been working on making small models more capable at agentic coding and work, because most people in the world don’t have the sort of hardware needed to run 3.8-27B, or even 35B-A3B or 9B dense, and I want to extend local agentic coding capability to less privileged users. This quant can be run…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-21 15:52 · r/LocalLLaMA
A better coder for the small-GPU/small-RAM crowd!