AINewsnow

How does Qwen 3.8 27B compare on low thinking mode to the older 3.6 models?

This story is from 2026-09-13. It is preserved in the archive; the latest stories are on the live feed.

Since we know Qwen 3.8 27B thinks quite long, but gives at least a good one-shot result where you can leave it to do everything on it own, how does it compare to the older series of models for very simple tasks where you don't want to think so long? The only fine-tune of Qwen 3.6 I genuinely enjoye…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-13 09:43 · r/LocalLLaMA
    How does Qwen 3.8 27B compare on low thinking mode to the older 3.6 models?

More stories

  1. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  2. Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis — r/LocalLLaMA
  3. US government website used Chinese model the FBI called "malicious" — Ars Technica AI
  4. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM
  5. Qwen Developers on X: "Qwen-Image 2.1 is going open source" — r/StableDiffusion
  6. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM
  7. Qwen Image 2.1 on Comfy: Coming Soon — r/StableDiffusion
  8. Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI

Get the daily brief of stories like this at 6:30 every morning →