AINewsnow

Qwen 3.8 27B reaching 225 t/s on 132k context on CMP 170HX

I reach around 225 t/s on a single CMP 170HX 40GB VRAM with Qwen 3.8 27b 132k context. I overclocked the GPU to NDIV 60 and I get 1.89 TB/s memory bandwidth and around 1500 mhz. If anyone is interested on a guide and they have a CMP GPU, i followed this guide from another post for optimizing the gp…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-23 21:48 · r/LocalLLM
    Qwen 3.8 27B reaching 225 t/s on 132k context on CMP 170HX

More stories

  1. Alibaba unveils new AI chip to challenge NVIDIA, plans Qwen models with up to 10 trillion parameters — Mint AI
  2. GPT image 2.5 vs Nano Banana pro vs Nano Banana 2 vs Qwen image 3 vs Seedream 5.0 pro — r/GeminiAI
  3. Help? — r/GeminiAI
  4. Alibaba's Qwen-Audio 3.1 Slashes Voice API Prices by up to 95% — AlphaSignal
  5. XiaomiMiMo/MiMo-V2.6-Pro-RL · Hugging Face — r/LocalLLaMA
  6. yandex/AliceAI-Foundation-80B-A3B-Base: Russian-developed competitor to Qwen 35B and DeepSeek V4 Flash — r/LocalLLaMA
  7. Alibaba Cloud Unveils Agentic Cloud Stack: AgentCore, Next-Gen CPFS and HPN 8.0 Pro — Pandaily
  8. I built a small local studio to try Qwen-Image-2.1 on my Mac — sharing in case you want to test it too — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →