AINewsnow

DeepSeek 4.1: MoE de 1,2 Trilhão com MLA v2 e DualPipe

A segunda quinzena de setembro de 2026 marca a consolidação de um novo ápice na engenharia global de modelos de inteligência artificial de fronteira. Em um anúncio simultâneo que impactou os centros de computação de alta performance em Hangzhou, Pequim, Vale do Silício e Londres, a DeepSeek oficial…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 1 report

  1. 2026-09-28 20:59 · DEV Community — Machine Learning
    DeepSeek 4.1: MoE de 1,2 Trilhão com MLA v2 e DualPipe

More stories

  1. PSA: Dual 3090 - Qwen Flash Next - 80tps/2k+ prefill — r/LocalLLM
  2. 85 GB DeepSeek-V4-Flash at ~3 tok/s on a 12 GB RTX 3060 + 64 GB DDR5 RAM - Overspill for FreeToken, inspired by Colibri — r/LocalLLaMA
  3. Worth going from Qwen3.8 27B to flash next or maaybe deepseek v4 flash? — r/LocalLLaMA
  4. Another "Harness matters" post (codex cli > pi and opencode) — r/LocalLLaMA
  5. How GLM5.3 Sparse Attention Affects HBM Memory Usage — SemiAnalysis
  6. Opus 5.5 (high) improves on Opus 5 (high) 3.5 → 3.8 on the Short-Story Creative Writing Benchmark, just behind Fable 5.1 (high) and Opus 5 (xhigh). — r/singularity
  7. Why are Gemini's answers outdated? — r/Bard
  8. Space is beautiful Pi coding agent plus Deepseek 41 — r/aivideo

Get the daily brief of stories like this at 6:30 every morning →