Best small model + setup together with qwen3.8-flash-next
hello, I currently use swift qwen3.8 gsq rco in iq3 xxs in strata. The model is only used for hermes agent. I have 64GB of ram + 32GB R9700 where the model runs on. But I also have an rx7800xt with 16GB. For hermes agent i want to run a small but intelligent enough model as an executioner model, wh…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-10 21:54 · r/LocalLLaMA
Best small model + setup together with qwen3.8-flash-next