AINewsnow

Meta Releases Llama 4 Scout: Frontier Open-Weights Mixture-of-Depths Architecture with Native 1M-Token Context Window

This story is from 2026-10-11. It is preserved in the archive; the latest stories are on the live feed.

Meta's Fundamental AI Research (FAIR) team has officially released Llama 4 Scout , the open-source community's first frontier-class foundation model built natively on dynamic Mixture-of-Depths (MoD) routing. Sporting a total parameter capacity of 105 billion with only 24 billion active parameters p…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-10-11 04:56 · DEV Community — AI
    Meta Releases Llama 4 Scout: Frontier Open-Weights Mixture-of-Depths Architecture with Native 1M-Token Context Window

More stories

  1. Qwen 3.6 35B A3B: 131K context + vision on 6GB VRAM — r/LocalLLaMA
  2. Tested Mellum2.1-12B-A2.5B on PI Coding Agent - surprisingly usable, but not great at one-shot projects — r/LocalLLaMA
  3. Running the uncensored Qwen3.8-27B (HauhauCS) on a 4090 at 262K context and ~130 tok/s — r/LocalLLaMA
  4. Java vllm-like framwork claims 90% of perfomance of llama.cpp on local inference on NVIDIA GPUs by compiling Java to CUDA and cuTile — r/LocalLLM
  5. Qwen3.8-Flash-Next (125B) at ~100 tok/s on an M5 Ultra Mac Studio with llama.cpp — r/LocalLLM
  6. Success with Qwen3.8 27B GSQ-RCO-IQ3_S on 16GB VRAM — r/LocalLLM
  7. Created laya : Now Introducing a new 800 Million Param physics-based typed decision model with 73k context and image support — r/LocalLLaMA
  8. Pancho The Llama (Pt I) - Mean People Suck. — r/comfyui

Get the daily brief of stories like this at 6:30 every morning →