2x RTX 3060s (12gb each) Qwen3.8-27B-exl3-4.0bpw 20-40 tokens per second ~170k context
Any tips to improve?
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-23 18:00 · r/LocalLLM
2x RTX 3060s (12gb each) Qwen3.8-27B-exl3-4.0bpw 20-40 tokens per second ~170k context