[Splash Engine] Qwen3.8-27B in native 8-bit at 37–55 tok/s on Apple Silicon: Extending Splash to Q8, 256k context scaling, and the "Reasoning Cliff"
Coverage of "[Splash Engine] Qwen3.8-27B in native 8-bit at 37–55 tok/s on Apple Silicon: Extending Splash to Q8, 256k context scaling, and the "Reasoning Cliff"" from 1 source, with a live timeline of who reported what and when.
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-21 12:34 · r/LocalLLM
[Splash Engine] Qwen3.8-27B in native 8-bit at 37–55 tok/s on Apple Silicon: Extending Splash to Q8, 256k context scaling, and the "Reasoning Cliff"
More stories
- iPhone owners can now submit claims in Apple’s $250 million Siri AI settlement — The Verge AI
- Apple Mac Studio (M5 Ultra) review: Huge AI and graphics power at a huge premium — Engadget
- Q&A with Mark Gurman on the iPhone Duo, breaking Apple news, Apple's AI-native devices, Tim Cook staying as executive chair, John Ternus, Johny Srouji, and more (Nilay Patel/The Verge) — Techmeme
- Can John Ternus find Apple’s next big thing? — The Verge AI
- Got an Android Phone? Google Thinks You’ll Probably Want a Googlebook Laptop — Wired AI
- Apple home hub device built around Siri AI may be almost ready — Engadget
- Gemini Joins the Hacker Club — Wall Street Journal Technology
- ComfyUI on Apple Silicon: no MLX, no fp8, 600-second kernel builds. So I built my own launcher — a personal project I'm sharing in case it helps someone. — r/comfyui
Get the daily brief of stories like this at 6:30 every morning →