What are some tricks to get Qwen 3.8 27B to run faster?
I have it running on an RTX 3090, 64K context, q5, xhigh reasoning, just as an always on personal assistant - reading emails, calendar, updating google docs/sheets, memory files, etc. The harness is vercel's eve framework, with a few patches. While the quality is pretty great, for 'live' chatting/t…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-03 05:47 · r/LocalLLM
What are some tricks to get Qwen 3.8 27B to run faster?