Cerebras Runs Alibaba's Qwen 3.8 27B at 1,850 Tokens per Second
This story is from 2026-09-11. It is preserved in the archive; the latest stories are on the live feed.
Alibaba's 27B dense multimodal model lands on Cerebras at roughly 1,800 tokens per second, with reasoning on by default and a 128K context on paid tiers.
Read the full story at AlphaSignal ↗
Timeline · 1 report
- 2026-09-11 20:09 · AlphaSignal
Cerebras Runs Alibaba's Qwen 3.8 27B at 1,850 Tokens per Second