AINewsnow

Meta AI Unveils Llama 3.2 Open‑Source Release, 70B Model

This story is from 2026-09-03. It is preserved in the archive; the latest stories are on the live feed.

Meta AI Unveils Llama 3.2 Open‑Source Release, 70B Model Meta AI announced the open‑source release of Llama 3.2 on May 28, 2024, delivering a 70‑billion‑parameter language model under a permissive license. The milestone marks the largest publicly available model from the company and signals a shift…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-03 12:34 · DEV Community — AI
    Meta AI Unveils Llama 3.2 Open‑Source Release, 70B Model

More stories

  1. Qwen3.8-Flash-Next-Heretic2-IQ4XS on Halogen Flash Server vs llama-server on Strix Halo: 2.3-7.7x prefill speedup with half the VRAM (+ vision works on BYO GGUF) — r/LocalLLM
  2. M2 Mac ultra128gb Qwen flash next — r/LocalLLM
  3. Multi-hour llama.cpp optimization experiments on Qwen MoE models, patches, benchmarks, and reproduction guides — r/LocalLLM
  4. CUDA: enable sparse fa for qwen4 by am17an · Pull Request #28770 · ggml-org/llama.cpp — r/LocalLLaMA
  5. focus-llama: a llama.cpp fork implementing Declarative Attention (arXiv:2609.02737) — r/LocalLLaMA
  6. I benchmarked 13 model/quant configs on a GPU with no tensor cores (Vega iGPU + Vulkan) and wrote it up as a measurement study — the quant encoding suffix matters more than you'd think — r/LocalLLM
  7. M1 Max 32GB, trying to run Qwen 3.8 27B at decent speeds and context — r/LocalLLaMA
  8. I turned an asymetric pair of Tesla V100s PCIe both (16 GB + 32 GB) into a surprisingly capable local LLM lab — 1.38k prompt tok/s, 40 decode tok/s with qwen3.8 27B Q6 and Q8... — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →