Qwen3.8-Flash-Next on Strata
Hey! 👋 I have released an official support for Strix Halo machines on Strata for Qwen3.8-Flash-Next. Currently numbers are the best on long context decode and ppts using typical Unsloth’s Q4 and GSQ-RCO model weights. Can go up to 1M context length without big speed loss. Currently support is mark…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-06 14:58 · r/LocalLLaMA
Qwen3.8-Flash-Next on Strata
More stories
- GPT-6 and Intelligent UI for everyone — OpenAI News
- Introducing Mistral Large 4 — Mistral AI News
- Introducing Claude Haiku 5.5 on AWS — AWS Machine Learning Blog
- Mistral Says Its New AI Model ‘Le Chonk’ Is the Best Open-Weight Offering Outside of China — Wired AI
- Sharing AI progress in mathematics — OpenAI News
- Introducing Playground: Create and play custom games — Google AI Blog
- OpenAI Decisions API now available on AI Gateway — Vercel Blog
- Everything announced at Microsoft's Surface Laptop Ultra event — The Verge AI
Get the daily brief of stories like this at 6:30 every morning →