Built a home server from an old PC with GPU upgrade. Qwen3.8 27B runs at ~30 tokens per second.
I needed a relatively simple but acceptable level of AI for working on one project. I didn't have any heavy requests, I just needed to give the AI access to the project files so it could search through them for bugs and stuff. I already had an old computer that I decided not to throw away and inste…
Read the full story at r/LocalLLaMA ↗
Timeline · 2 reports
- 2026-09-20 10:42 · r/LocalLLM
Qwen3.8 27b - what is realistic tokens per second with 16 GB VRAM and 32 GB RAM? - 2026-09-19 08:09 · r/LocalLLaMA
Built a home server from an old PC with GPU upgrade. Qwen3.8 27B runs at ~30 tokens per second.