AINewsnow

I Ran 3 Open-Weight LLMs Head-to-Head on a 24GB Mac — One Was 3x Faster

This story is from 2026-09-04. It is preserved in the archive; the latest stories are on the live feed.

I ran three Apache 2.0 open-weight LLMs — OpenAI's gpt-oss-20B, Alibaba's Qwen3-14B, and Mistral's Mistral-Small-24B — on the same Apple M2 24GB machine, through the same Ollama 0.12.x runtime, on the same four benchmark tasks. The fastest model is the one I didn't expect to win. The most accurate…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-04 00:07 · DEV Community — AI
    I Ran 3 Open-Weight LLMs Head-to-Head on a 24GB Mac — One Was 3x Faster

More stories

  1. Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
  2. I gave 6 different AIs the same 5 questions — r/AI_Agents
  3. Local LLM on iPhone 18 is impressive — r/LocalLLM
  4. Has anyone had this happen? — r/OpenAI
  5. [AINews] OpenAI reports Navier-Stokes singularity find in 88 hours using Astra-next, roughly 10,000 agents and 130B tokens (>$40M), a contender for second ever Millennium Prize awarded — Latent Space
  6. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  7. Microsoft and OpenAI Workers Worry About ‘Largest Theft of Labor’ in History — New York Times Technology
  8. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI

Get the daily brief of stories like this at 6:30 every morning →