AINewsnow

Did anyone do a full bench of e.g. Qwen Flash Next IQ4 and Qwen 27b FP8? Here are some

edit: had to change formatting as I’m on mobile noe and the table broke.. I let Codex do a quick eval on Qwen 3.8 Flash Next IQ4_XS (served via vLLM + r9v) and Qwen 3.8 27B FP8 (served via vLLM + Radiance) This was not the full eval. I only ran a reduced quick test because the full version would ta…

Read the full story at r/LocalLLaMA ↗

Timeline · 6 reports

  1. 2026-09-26 21:11 · r/huggingface
    Abbiamo inserito 100 informazioni nella tabella engrammatica di Qwen 3.8 Flash Next e abbiamo creato un sito web per illustrarle.
  2. 2026-09-26 19:54 · r/LocalLLaMA
    Qwen 3.8 flash next is based on Qwen 4 architecture, if the announced Qwen 4 27b is also the same architecture with n-grams does it mean I can actually have faster inference on a single 3090 without tweaking much?
  3. 2026-09-26 13:02 · r/LocalLLM
    We have implanted 100 facts into the engram table of Qwen 3.8 Flash Next, and we have now created a website to explain it.
  4. 2026-09-26 08:52 · r/LocalLLaMA
    Best current Qwen Flash Next Q4-ish? + worth using?
  5. 2026-09-26 03:42 · r/LocalLLM
    PSA: Dual 3090 - Qwen Flash Next - 80tps/2k+ prefill
  6. 2026-09-25 06:32 · r/LocalLLaMA
    Did anyone do a full bench of e.g. Qwen Flash Next IQ4 and Qwen 27b FP8? Here are some

More stories

  1. Another "Harness matters" post (codex cli > pi and opencode) — r/LocalLLaMA
  2. Which provider actually wins on pure affordability right now for gemma qwen gpt oss and deepseek under one roof — r/AI_Agents
  3. When is the next generation of "B tier" models releasing? — r/LocalLLaMA
  4. One key for claude, gpt, gemini, and deepseek in my coding tools — r/ChatGPTCoding
  5. viggle-turbo isn't just faster - for most prompts, it's just as good — r/StableDiffusion
  6. Swift 1.5 Qwen3.8 27b (A must-have for low thinking!) — r/LocalLLaMA
  7. Adding logit penalty for "wait", "maybe" and "perhaps" to Qwen models improves their accuracy — r/LocalLLaMA
  8. I built a Fooocus-style local studio for Qwen-Image 2.1: masks, annotations, outpainting, OpenPose poses and sketches as references. Looking for honest criticism — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →