Finally got qwen 3.8 q4 k_M 27b runnin slick 40tkps load at 240k inf 4qnl for stable agentic coding on a 7900xtx. I have been finally able to give hermes a task and come back to results. Built a panel to manage my inference servers from hermes.
https://github.com/W61k3r/LexiPanel for those interested in the panel Screenshot of heavy load. I did offload vision to a spare 2060 instead of local ram. The control panel was generated with qwen3.8-27b davidau q4k_m max mtp. Made the panel from scratch over a week, basically just building for spe…
Read the full story at r/LocalLLaMA ↗
Timeline · 3 reports
- 2026-09-20 03:53 · r/LocalLLM
Finally got Qwen 3.8 Next running on my v100 6gpu setup (TP2 PP3) - 2026-09-20 03:44 · r/LocalLLaMA
Finally got Qwen 3.8 Next running on my v100 6gpu setup (TP2 PP3) - 2026-09-19 19:44 · r/LocalLLaMA
Finally got qwen 3.8 q4 k_M 27b runnin slick 40tkps load at 240k inf 4qnl for stable agentic coding on a 7900xtx. I have been finally able to give hermes a task and come back to results. Built a panel to manage my inference servers from hermes.