im making qwen3.8:27b into a MoE and its not half bad...
I have a rather low grade gpu (8 gig laptop GPU) and ive been wanting to run qwen3.8:27b but at the time I simply couldn't. I know that there are some 1-2 bit quants but they dont seem to good. so I made 2 things 1, a custom inference engine made especially for MoE architectures 2, a program that s…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-27 17:22 · r/LocalLLM
im making qwen3.8:27b into a MoE and its not half bad...