AINewsnow

Local Qwen3.8-Flash on a DGX Spark vs Claude Opus 5.5: 21 graded tasks, 3 harnesses. Matches Opus on everyday coding, 70% on hard tasks, and the harness settings matter more than you'd think

Human written: Been toying with local models for a while, to various degrees of success. The primary incentive to look into the Qwen 3.8 family was because I stopped the Claude Max subscription, and I'm running out of tokens on a somewhat regular basis. Had the $100/mo Max for several months as I w…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-08 00:36 · r/LocalLLM
    Local Qwen3.8-Flash on a DGX Spark vs Claude Opus 5.5: 21 graded tasks, 3 harnesses. Matches Opus on everyday coding, 70% on hard tasks, and the harness settings matter more than you'd think

More stories

  1. Qwen 3.8 Flash-Next on 64GB RAM + 16GB VRAM, worth it over a fully-loaded 3.6 35B-A3B? — r/LocalLLaMA
  2. What to run overnight and such when llm idle? — r/LocalLLM
  3. 113 Decision Models in 3 Weeks (Mostly Qwen and Gemma Fine-Tunes): A Deep Dive — r/AI_Agents
  4. Ultimate web scraper — r/AI_Agents
  5. Switching from Claude Code to local Qwen for Android dev — can smaller local models keep up? — r/LocalLLM
  6. I built a root-cause tool with Claude Code, then built a validator to catch the LLM inside it when it's confidently wrong. Honest numbers: 60% / 20% — r/AI_Agents
  7. "I don't want an oligopoly": New open-weight AI models mount a comeback against China — Axios AI+
  8. GPT-6 and Intelligent UI for everyone — OpenAI News

Get the daily brief of stories like this at 6:30 every morning →