Deepseek V4 Flash 0731
This story is from 2026-09-10. It is preserved in the archive; the latest stories are on the live feed.
I have three DGX Spark. On one of them I run 3.8 27B and a few vision models. On the other two I generally have been running Qwen 3.8 Next / Flash with a decent recipe getting 39.8 t/s mean decode across controlled and uncontrolled coding both thinking+answer. The "final code" decode is roughly 74…
Read the full story at r/LocalLLM ↗
Timeline · 27 reports
- 2026-09-12 19:09 · r/huggingface
We’re testing DeepSeek V4.1 flash bs GLM 5.3 flash. V4.1 is free to use. - 2026-09-12 17:12 · r/LocalLLM
DeepSeek V4.1 Flash running locally on 8× A40 — ~40 tok/s Q2_K, ~32 tok/s Q4_K_M - 2026-09-12 05:56 · Latent Space
[AINews] DeepSeek v4.1-Flash: 763B-P8B-D16B novel causal Encoder–Decoder architecture with vision marks the Return of the Whale - 2026-09-11 22:24 · Unite.AI
Baseten Adds DeepSeek-V4.1-Flash to Model APIs With 1M-Token Context - 2026-09-11 14:56 · r/huggingface
DeepSeek-V4.1-Flash: GGUF + 4.75bpw EXL3 are out, looking for devs with 4× DGX Sparks to help validate the EXL3 TP4 recipe - 2026-09-11 03:09 · Pandaily
Cambricon Day-0 Adapts DeepSeek-V4.1-Flash on vLLM Stack - 2026-09-10 23:45 · SiliconANGLE AI
DeepSeek releases V4.1-Flash, says it outperforms flagship V4-Pro - 2026-09-10 20:00 · r/LocalLLaMA
CPU Only Experimental Sloppy Deepseek V4.1 Flash - 2026-09-10 19:45 · r/LocalLLaMA
Livebench added Deepseek v4.1 flash - 2026-09-10 19:05 · r/LocalLLM
DeepSeek V4.1 Flash is 510 GB but only about 150 of it has to be in memory. I read the shard headers and made a fit checker. - 2026-09-10 16:05 · r/huggingface
DeepSeek V4.1 Flash is available in HuggingChat - 2026-09-10 16:05 · r/LocalLLaMA
DeepSeek V4.1 Flash is available in HuggingChat - 2026-09-10 11:23 · r/LocalLLaMA
Deepseek V4.1 Flash Release Video [Made with Deepseek V4.1 Flash] - 2026-09-10 10:01 · r/LocalLLaMA
guide to using reasoning_effort on deepseek v4.1 flash - 2026-09-10 08:37 · r/LocalLLaMA
DeepSeek-V4.1-Flash surprised .... - 2026-09-10 08:27 · r/LocalLLaMA
Deepseek V4.1 Flash is 748B, not 552B - 2026-09-10 07:56 · r/singularity
DeepSeek V4.1 Flash is getting surprisingly close to GPT-5.6 Sol territory, while being absurdly cheap - 2026-09-10 07:56 · r/Bard
DeepSeek just dropped V4.1 Flash, anyone tried it? - 2026-09-10 07:44 · TechNode
DeepSeek releases Harness 0.1.5 with V4.1 Flash support, file uploads and sidebar previews - 2026-09-10 07:31 · r/LocalLLM
DeepSeek V4.1 Flash in 3 charts: vs its predecessor, a top open-weight rival, and Claude Opus 5 - 2026-09-10 07:02 · r/LocalLLM
DeepSeek releases DeepSeek-V4.1-Flash! - 2026-09-10 07:00 · r/LocalLLaMA
Deepseek v4.1 flash finally has engrams, what do you expect from 4.1 pro? - 2026-09-10 06:54 · r/LocalLLaMA
DeepSeek V4-1 Flash is out - 2026-09-10 06:48 · r/singularity
DeepSeek v4.1 Flash Benchmarks - 2026-09-10 06:27 · r/LocalLLaMA
DeepSeek V4.1 Flash: Stronger, Faster, More Accessible - 2026-09-10 06:08 · r/LocalLLM
DeepSeek-V4.1-Flash is out - 2026-09-10 04:13 · r/LocalLLM
Deepseek V4 Flash 0731