AINewsnow

Daily driving Qwen 3.8 Flash instead of Claude.

Even 6 months ago, I wouldn't have thought I would be replacing Claude with a local LLM. But here we are. All these PRs are done by Qwen 3.8 (mostly flash, some 27b) running locally on my laptop (strix halo 138gb). And its perfectly acceptable speed (1200+ prefil, 40t/s decode) and Opus 4.8 level q…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-04 18:47 · r/LocalLLM
    Daily driving Qwen 3.8 Flash instead of Claude.

More stories

  1. How I use a local Qwen 27B for real work on home hardware (unscripted workflow demo) — r/LocalLLM
  2. Claude and Grok built me a local monitoring setup for my AI box: two dashboards, one for the machine, one for model training — r/LocalLLM
  3. I put Gemma 4 26B and Qwen 3.8 27B (2x3090) in charge of a club in Championship Manager 97/98, against Claude, Grok and DeepSeek. It's running now. — r/LocalLLM
  4. I made a simple gui for claude code, usable with local llms — r/LocalLLM
  5. Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context. — r/LocalLLM
  6. Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4 — r/LocalLLaMA
  7. The Museum of Lost Things | Short Film by Claude (Minimax H3) NO user input. — r/ClaudeAI
  8. Add secure Web Search to Claude Desktop with Amazon Bedrock AgentCore — AWS Machine Learning Blog

Get the daily brief of stories like this at 6:30 every morning →