AINewsnow

The 27B Class: August's Small Open-Weight Models, Compared

This story is from 2026-08-25. It is preserved in the archive; the latest stories are on the live feed.

Originally published on AI Tech Connect . What you need to know Three models, one shelf. Qwen3.8-27B, NVIDIA's Nemotron 3.5 Lightning and Meta's Muse Glimmer 30B all landed within three weeks in August 2026, all in the size class that fits a single 24–32GB machine. Same size, almost nothing else in…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 1 report

  1. 2026-08-25 11:31 · DEV Community — Machine Learning
    The 27B Class: August's Small Open-Weight Models, Compared

More stories

  1. Which models you run on your Nvidia v100? — r/LocalLLM
  2. Benchmarked llama.cpp vs llamafile vs LM Studio vs Ollama on 3 machines (Metal/Vulkan/CUDA). Throughput is almost identical until you change how they're built. — r/LocalLLM
  3. Am I right in thinking llama.cpp is the only show in town for mixed (Nvidia) GPUs? — r/LocalLLM
  4. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  5. Amid growing AI fears, King Charles meets with industry leaders in Scotland — NPR Technology
  6. Qwen3.8-Flash-Next-Heretic2-IQ4XS on Halogen Flash Server vs llama-server on Strix Halo: 2.3-7.7x prefill speedup with half the VRAM (+ vision works on BYO GGUF) — r/LocalLLM
  7. Meta Launches Muse Mac App With File, Messages, and Calendar Access — Unite.AI
  8. King Charles to press Nvidia, OpenAI, Anthropic leaders on AI safety at summit — CNBC Technology

Get the daily brief of stories like this at 6:30 every morning →