Local model choice: why is (not) that hard
TL;DR - ≥32 GB VRAM → Qwen3.8-27B , Q4 (NVFP4 / Q4_K_XL / IQ4_XS) - Otherwise → Qwen3.8-Flash-Next on Strata , Q4 or IQ3 depending on your total RAM + VRAM For the past few weeks I've been deep into local models: testing everything I could get my hands on, reading up on which hardware does what, an…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-02 21:26 · r/LocalLLM
Local model choice: why is (not) that hard