Qwen3.8-27B Q4_K_M on one RTX 3090 + OpenCode: throughput, four coding tasks, and a reasoning-budget failure
I put an old RTX 3090 to work as a local coding agent with Qwen3.8-27B, llama.cpp, and OpenCode. Here are the setup and results, including what failed. This is a summary of my own blog post, linked below. Setup RTX 3090 24GB, Ryzen 7 5800X, 64GB RAM, Ubuntu. Qwen3.8-27B Q4_K_M weights (~16.8GB), al…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-01 20:04 · r/LocalLLaMA
Qwen3.8-27B Q4_K_M on one RTX 3090 + OpenCode: throughput, four coding tasks, and a reasoning-budget failure