AINewsnow

Qwen3.8-27B: How a 3:1 Hybrid Attention Ratio Lets a 27B Model Punch Above Its Weight

This story is from 2026-08-24. It is preserved in the archive; the latest stories are on the live feed.

Qwen3.8-27B: How a 3:1 Hybrid Attention Ratio Lets a 27B Model Punch Above Its Weight Alibaba's Tongyi Lab released Qwen3.8-27B on August 14, 2026 — a 27.78-billion-parameter dense multimodal model that makes a specific architectural bet: replace three out of every four attention layers with a line…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 1 report

  1. 2026-08-24 16:21 · DEV Community — Machine Learning
    Qwen3.8-27B: How a 3:1 Hybrid Attention Ratio Lets a 27B Model Punch Above Its Weight

More stories

  1. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  2. Alibaba open-sources medical AI model that can detect cancer and nearly 150 conditions — South China Morning Post Tech
  3. Alibaba's Damo Academy open sources RADAR, a medical vision-language model it says can read CT scans and identify ~150 abdominal conditions, including cancers (Ann Cao/South China Morning Post) — Techmeme
  4. Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis — r/LocalLLM
  5. 10 hours left fo Qwen Image 2.1 Public Open Source Release — r/StableDiffusion
  6. US government website used Chinese model the FBI called "malicious" — Ars Technica AI
  7. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM
  8. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →