AINewsnow

Ling 3.0 Tiny makes an amazing auxillery model for Hermes (Qwen 3.8 27B as the primary model)

This story is from 2026-08-20. It is preserved in the archive; the latest stories are on the live feed.

Just got the setup dialed in yesterday. Getting really good results and is making 3.8 usage feel faster in hermes. I got Qwen to specifically use Ling tiny for simple tasks like context compression and summerization tasks. (basically anything that is not intellegence critical) Ling has like 5x fast…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-08-20 00:13 · r/LocalLLaMA
    Ling 3.0 Tiny makes an amazing auxillery model for Hermes (Qwen 3.8 27B as the primary model)

More stories

  1. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  2. US government website used Chinese model the FBI called "malicious" — Ars Technica AI
  3. Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis — r/LocalLLM
  4. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM
  5. Qwen Developers on X: "Qwen-Image 2.1 is going open source" — r/StableDiffusion
  6. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM
  7. Qwen Image 2.1 Examples — r/StableDiffusion
  8. Qwen Image 2.1 on Comfy: Coming Soon — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →