AINewsnow

Which LLM is best for coding/agents if I have dual R9700 GPUs?

I’m looking for recommendations for the best local LLM for coding and agentic tasks. I’m running dual AMD R9700 32GB GPUs (64GB VRAM total). It is linux llama.cpp server. I’ve tried Qwen 3.8 27B Q8 with a 200K context, and it worked reasonably well on my large app, but it still made some coding mis…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-01 14:50 · r/LocalLLM
    Which LLM is best for coding/agents if I have dual R9700 GPUs?

More stories

  1. Qwen flash next on 12+16gb vram, and 32gb ram viable? — r/LocalLLM
  2. Sonnet 5.5 orchestrated a local Qwen 3.8 27B! — r/ClaudeAI
  3. add GLM-5.3-Flash (GLM5-Next) support by timkhronos · Pull Request #27773 · ggml-org/llama.cpp — r/LocalLLaMA
  4. Used Opus 5.5 to optimize llama.cpp inference for Swift Qwen 3.8 27B Q6_K on RTX 5090 - decode 143 tok/s prefill 2840 tok/s — r/LocalLLM
  5. Qwen 3.8 27B on a single 3090: 114 min solo, 43 min as a worker under a GPT 6.1 SOL orchestrator — r/LocalLLM
  6. Qwen 3.8 27B Q4 with 100K context on a 16 GB RX 7800 XT guide — r/LocalLLaMA
  7. Qwen 3.8 27B on a 3090 with a Sonnet 5.5 as a planner: 2.7x cheaper, real numbers — r/LocalLLM
  8. Qwen 3.8 is a workhorse — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →