Two node BC250 cluster comparison of Qwen3.6 vs Qwen 3.8
TLDR: Qwen3.6-35B-A3B made a single game in a little over a minute and Qwen3.8-27B made a much better game but took close to an hour. Upon request I =tn a one-shot creative test tonight on my two-node AMD BC-250 llama.cpp cluster (Vulkan + RPC, MTP speculative decoding on, 115k context). Same byte-…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-19 04:00 · r/LocalLLM
Two node BC250 cluster comparison of Qwen3.6 vs Qwen 3.8