Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model
Mistral Large 4, nicknamed Le Chonk, went into public preview today. 1.05T total parameters, 49B active per token, 1M context, native image input. Trained from scratch on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's own European datacenters. Specs (from the model docs) 1.05T total params, 49B act…
Read the full story at r/machinelearningnews ↗
Timeline · 4 reports
- 2026-10-07 14:44 · r/huggingface
Mistral Large 4 "Le Chonk": The 1-Trillion Parameter Open-Source Monster That Changes Everything? - 2026-10-07 00:12 · Matthew Berman
Mistral is BACK! (Le Chonk) - 2026-10-06 20:18 · Simon Willison's Weblog
Introducing Mistral Large 4: Le chonk - 2026-10-06 17:54 · r/machinelearningnews
Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model