AINewsnow

MicroLLM Lab: 7 Open Models, Side-by-Side Task Accuracy

This story is from 2026-09-29. It is preserved in the archive; the latest stories are on the live feed.

Hacker News posted MicroLLM Lab yesterday — an open-source benchmark of 7 sub-3B models across 12 real-world tasks. I re-ran the suite on my own hardware to verify. The Models Llama 3.2 1B — Meta, general purpose Phi-3.5 Mini 3.8B — Microsoft, reasoning focus Qwen 2.5 1.5B — Alibaba, multilingual G…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-29 11:09 · DEV Community — AI
    MicroLLM Lab: 7 Open Models, Side-by-Side Task Accuracy

More stories

  1. We have implanted 100 facts into the engram table of Qwen 3.8 Flash Next, and we have now created a website to explain it. — r/LocalLLM
  2. Qwen 3.8 27B on a 3090 with a Sonnet 5.5 as a planner: 2.7x cheaper, real numbers — r/LocalLLM
  3. Qwen 3.8 is a workhorse — r/LocalLLaMA
  4. vulkan: fuse qwen4exp's SCALE -> SIGMOID -> SCALE -> hc_post chain by fxgsell · Pull Request #29520 · ggml-org/llama.cpp — r/LocalLLaMA
  5. Adding logit penalty for "wait", "maybe" and "perhaps" to Qwen models improves their accuracy — r/LocalLLaMA
  6. Layer Extract & Layer Remove Loras For Qwen Image 2.1 — r/StableDiffusion
  7. Qwen 3.8 27B vs Qwen 3.8 Flash Next and time to complete a coding task. — r/LocalLLaMA
  8. I built Slopus, a free, open-source desktop app for generating and editing AI videos locally (Minimax H3) — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →