Qwen3.8-Flash-Next Q4 vs Qwen3.8 27B Q5 on single R9700 (32GB) + 64GB RAM: 2x128k context, almost similar performance
Hey everyone, Spent the last week setting up a local rig for agent work (Hermes Agent: a cloud model plans and reviews, local models do the work) and comparing Qwen3.8-27B unsloth Q5_K_XL with Flash-Next Q4 on a single R9700 with 64 GB RAM. Took a lot of trial and error, so sharing what worked. Use…
Read the full story at r/LocalLLM ↗