AINewsnow

Mistral releases Mistral Large 4, dubbed "le Chonk", a 1T-parameter open-weight model for general agentic capabilities, trained on 4,000 Grace Blackwell GPUs (Sabrina Ortiz/The Deep View)

Sabrina Ortiz / The Deep View : Mistral releases Mistral Large 4, dubbed le Chonk , a 1T-parameter open-weight model for general agentic capabilities, trained on 4,000 Grace Blackwell GPUs Most flagship models dominating the market are from frontier labs and are closed models. French AI lab Mistral…

Read the full story at Techmeme ↗

Timeline · 8 reports

  1. 2026-10-06 13:15 · Wired AI
    Mistral Says Its New AI Model ‘Le Chonk’ Is the Best Open-Weight Offering Outside of China
  2. 2026-10-06 13:08 · Techmeme
    Mistral releases Mistral Large 4, dubbed "le Chonk", a 1T-parameter open-weight model for general agentic capabilities, trained on 4,000 Grace Blackwell GPUs (Sabrina Ortiz/The Deep View)
  3. 2026-10-06 06:35 · r/LocalLLaMA
    Reflection’s Beam 501B-A23B - open-weight model release this month
  4. 2026-10-06 00:46 · SiliconANGLE AI
    Reflection AI debuts open-source Beam model with 501B parameters
  5. 2026-10-05 21:10 · r/machinelearningnews
    Reflection AI introduces Beam: a 501B open-weight MoE with 23B active parameters, 1M context and Apache 2.0 weights coming this month
  6. 2026-10-05 21:04 · MarkTechPost
    Reflection AI Introduces Beam: A 501B Open-Weight MoE Model With 23B Active Parameters for Coding and Agentic Workloads
  7. 2026-10-05 19:46 · Unite.AI
    Reflection AI Unveils Beam, a 501B-Parameter Open-Weight Model
  8. 2026-10-05 05:33 · r/LocalLLaMA
    Reflection AI Is About to Release a US Open-Weight Model to Take On DeepSeek and Qwen

More stories

  1. I let 5 AI models fight a world war. DeepSeek betrayed Claude and nuked it four times. Mistral nuked itself. — r/AI_Agents
  2. A benchmark for LLMs playing Civilization V. GLM-5.3 is ahead of Opus-5.5, and Qwen-3.8-27B holds up surprisingly well. — r/LocalLLaMA
  3. Built a quick, sub-15ms Rust CLI/TUI to pack repos into prompts without burning 40k tokens on lockfiles and junk — r/LocalLLaMA
  4. Built a gateway so you can call DeepSeek, Qwen, Kimi, GLM, MiniMax with one key — USD billing, OpenAI-compatible — r/LocalLLM
  5. Introducing Mistral Large 4 — Mistral AI News
  6. Introducing GLM 5.3 on Amazon Bedrock — AWS Machine Learning Blog
  7. can i run qwen flash next with these specs, or am i out of luck? — r/LocalLLM
  8. The Story of Qwen: Alibaba's AI Models From 7B to 2.4T — MarkTechPost

Get the daily brief of stories like this at 6:30 every morning →