AINewsnow

DeepSeek-V4-Flash-Vision Q8 vs Qwen3.8-Flash-Next Q8

This story is from 2026-09-06. It is preserved in the archive; the latest stories are on the live feed.

I'm using DS-V4-Flash-Vision with Q8_K_XL quantization locally as my everyday engine, and for some time now I've been doing a lot of comparisons with Qwen3.8-Flash-Next, also with Q8_K_XL quantization. It took me quite a while to get Q3.8FN to work reasonably well, and here are my observations. My…

Read the full story at r/LocalLLaMA ↗

Timeline · 9 reports

  1. 2026-09-09 14:20 · r/singularity
    Deepseek v4.1 Flash reaches 98% of Astra’s score at 1.4% of cost on OpenDesign Arena
  2. 2026-09-09 10:50 · r/LocalLLaMA
    DeepSeek-V4-Flash-Vision-Exp (285B MoE) on 10-12x RTX 3090 — spec decoding, vision
  3. 2026-09-09 09:54 · r/LocalLLM
    DeepSeek V4 Flash Vision UD GGUF (mmproj) not loading in Unsloth Studio/Desktop?
  4. 2026-09-08 22:13 · r/ArtificialInteligence
    Tested DeepSeek V4 vs V4.1 Flash Vision Beta in 5 visual tests
  5. 2026-09-08 22:12 · r/LocalLLM
    V4.1 Flash Vision Beta is much more reliable AND cheaper
  6. 2026-09-08 01:22 · r/LocalLLM
    Qwen3.8 Flash Next vs Deepseek v4 0731 vs GLM 5.3 - A clear winner?
  7. 2026-09-07 18:27 · r/LocalLLaMA
    DeepSeek-V4-Flash-Vision-Exp is amazing at creating game worlds!
  8. 2026-09-07 02:31 · r/LocalLLM
    Unsloth template for DeepSeek-V4-Flash-Vision-Exp-GGUF
  9. 2026-09-06 20:15 · r/LocalLLaMA
    DeepSeek-V4-Flash-Vision Q8 vs Qwen3.8-Flash-Next Q8

More stories

  1. The Sequence Learning Loop - Issue 934: Understanding DeepSeek V4.1 Flash, DeepMind’s AlphaGenome Atlas and Muse — TheSequence
  2. DeepSeek’s Insane New Architecture — Two Minute Papers
  3. Jina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding for Low-Budget GPUs — MarkTechPost
  4. Gemini 4 is good enough - JUST RELEASE IT — r/GeminiAI
  5. Engrams Embedding Entendre: Codesign for Efficient DRAM/SSD Offloading — SemiAnalysis
  6. Coming soon...... Optimized for DEEPSEEK Flash.... Though model Agnostic.... message me to test.... cem888.ai — r/huggingface
  7. Laguna s 2.1 political censorship question — r/LocalLLM
  8. Bro why is the chat naming glitching and naming everything in chinese — r/ChatGPT

Get the daily brief of stories like this at 6:30 every morning →