AINewsnow

DeepSeek just dropped V4.1 Flash, anyone tried it?

This story is from 2026-09-10. It is preserved in the archive; the latest stories are on the live feed.

DeepSeek quietly dropped V4.1 Flash today. 552B params, MoE architecture, only 8B active params on input, 16B on output. KV cache shrank to one-437th of their original V1 model, which apparently makes agent workloads way more affordable. Benchmarks show it beating V4 Pro across most agentic tasks,…

Read the full story at r/Bard ↗

Timeline · 1 report

  1. 2026-09-10 07:56 · r/Bard
    DeepSeek just dropped V4.1 Flash, anyone tried it?

More stories

  1. Cactus Needle 3: A Sliceable 8-29MB Automation Foundation Model That Matches DeepSeek v4 Flash — r/LocalLLaMA
  2. DeepSeek’s Insane New Architecture — Two Minute Papers
  3. Jina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding for Low-Budget GPUs — MarkTechPost
  4. Gemini 4 is good enough - JUST RELEASE IT — r/GeminiAI
  5. Own 1 dashboard for ChatGPT, Gemini, Claude, and more for only $54.97 — Mashable AI
  6. I enjoyed the daily HF papers today — r/LocalLLaMA
  7. Engrams Embedding Entendre: Codesign for Efficient DRAM/SSD Offloading — SemiAnalysis
  8. Coming soon...... Optimized for DEEPSEEK Flash.... Though model Agnostic.... message me to test.... cem888.ai — r/huggingface

Get the daily brief of stories like this at 6:30 every morning →