DeepSeek-V4-Flash-Vision Q8 vs Qwen3.8-Flash-Next Q8
This story is from 2026-09-06. It is preserved in the archive; the latest stories are on the live feed.
I'm using DS-V4-Flash-Vision with Q8_K_XL quantization locally as my everyday engine, and for some time now I've been doing a lot of comparisons with Qwen3.8-Flash-Next, also with Q8_K_XL quantization. It took me quite a while to get Q3.8FN to work reasonably well, and here are my observations. My…
Read the full story at r/LocalLLaMA ↗
Timeline · 9 reports
- 2026-09-09 14:20 · r/singularity
Deepseek v4.1 Flash reaches 98% of Astra’s score at 1.4% of cost on OpenDesign Arena - 2026-09-09 10:50 · r/LocalLLaMA
DeepSeek-V4-Flash-Vision-Exp (285B MoE) on 10-12x RTX 3090 — spec decoding, vision - 2026-09-09 09:54 · r/LocalLLM
DeepSeek V4 Flash Vision UD GGUF (mmproj) not loading in Unsloth Studio/Desktop? - 2026-09-08 22:13 · r/ArtificialInteligence
Tested DeepSeek V4 vs V4.1 Flash Vision Beta in 5 visual tests - 2026-09-08 22:12 · r/LocalLLM
V4.1 Flash Vision Beta is much more reliable AND cheaper - 2026-09-08 01:22 · r/LocalLLM
Qwen3.8 Flash Next vs Deepseek v4 0731 vs GLM 5.3 - A clear winner? - 2026-09-07 18:27 · r/LocalLLaMA
DeepSeek-V4-Flash-Vision-Exp is amazing at creating game worlds! - 2026-09-07 02:31 · r/LocalLLM
Unsloth template for DeepSeek-V4-Flash-Vision-Exp-GGUF - 2026-09-06 20:15 · r/LocalLLaMA
DeepSeek-V4-Flash-Vision Q8 vs Qwen3.8-Flash-Next Q8