AINewsnow

Splash vs oMLX on the same uncensored Qwen3.8-27B on M3 MAX, 48 GB MBP

Saw all the Splash hype last week and wanted to know if it holds up on an older M3 Max, not the M5 Pro from Inco's charts. I've been running the OrcaRouter abliterated Qwen3.8-27B on oMLX with MTP on, getting low-to-mid 30s tok/s on normal writing, which I was pretty happy with. https://preview.red…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-02 00:06 · r/LocalLLM
    Splash vs oMLX on the same uncensored Qwen3.8-27B on M3 MAX, 48 GB MBP

More stories

  1. Bring near-Astra intelligence to everyday work with GPT-6.1 Sol on Amazon Bedrock — AWS Machine Learning Blog
  2. Gemini 4 Argon: our next era of frontier intelligence — Google Gemini Blog
  3. Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs — Allen Institute for AI (Ai2)
  4. Google Releases New Gemini Model With Guardrails Amid A.I. Safety Debate — New York Times Technology
  5. OpenAI scraps release of its latest AI model over safety concerns — France 24 — Artificial Intelligence
  6. Introducing GPT-6.1 Sol — OpenAI News
  7. Google tests its plan for AI data centers in space with Project Suncatcher — Scientific American
  8. OpenAI’s Dots Are Always-On AI Agents—and Its Answer to Meta’s Muse — Wired AI

Get the daily brief of stories like this at 6:30 every morning →