AINewsnow

[Splash Engine] Qwen3.8-27B in native 8-bit at 37–55 tok/s on Apple Silicon: Extending Splash to Q8, 256k context scaling, and the "Reasoning Cliff"

Coverage of "[Splash Engine] Qwen3.8-27B in native 8-bit at 37–55 tok/s on Apple Silicon: Extending Splash to Q8, 256k context scaling, and the "Reasoning Cliff"" from 1 source, with a live timeline of who reported what and when.

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-21 12:34 · r/LocalLLM
    [Splash Engine] Qwen3.8-27B in native 8-bit at 37–55 tok/s on Apple Silicon: Extending Splash to Q8, 256k context scaling, and the "Reasoning Cliff"

More stories

  1. iPhone owners can now submit claims in Apple’s $250 million Siri AI settlement — The Verge AI
  2. Apple Mac Studio (M5 Ultra) review: Huge AI and graphics power at a huge premium — Engadget
  3. Q&A with Mark Gurman on the iPhone Duo, breaking Apple news, Apple's AI-native devices, Tim Cook staying as executive chair, John Ternus, Johny Srouji, and more (Nilay Patel/The Verge) — Techmeme
  4. Can John Ternus find Apple’s next big thing? — The Verge AI
  5. Got an Android Phone? Google Thinks You’ll Probably Want a Googlebook Laptop — Wired AI
  6. Apple home hub device built around Siri AI may be almost ready — Engadget
  7. Gemini Joins the Hacker Club — Wall Street Journal Technology
  8. ComfyUI on Apple Silicon: no MLX, no fp8, 600-second kernel builds. So I built my own launcher — a personal project I'm sharing in case it helps someone. — r/comfyui

Get the daily brief of stories like this at 6:30 every morning →