I Ran 3 Open-Weight LLMs Head-to-Head on a 24GB Mac — One Was 3x Faster
This story is from 2026-09-04. It is preserved in the archive; the latest stories are on the live feed.
I ran three Apache 2.0 open-weight LLMs — OpenAI's gpt-oss-20B, Alibaba's Qwen3-14B, and Mistral's Mistral-Small-24B — on the same Apple M2 24GB machine, through the same Ollama 0.12.x runtime, on the same four benchmark tasks. The fastest model is the one I didn't expect to win. The most accurate…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-04 00:07 · DEV Community — AI
I Ran 3 Open-Weight LLMs Head-to-Head on a 24GB Mac — One Was 3x Faster