AINewsnow

5070 ti + 64 ram Qwen 3.8 27b

hey everyone, I am getting like 33-40 token speed at 65K Context, do you think this is good numbers? I usually use Qwen 3.8 27b Q4, I try to get better results by edit the bat files assisted by Claude, but I can't go above 40 t/s using LLM CCP, LM Studio.

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-23 16:29 · r/LocalLLM
    5070 ti + 64 ram Qwen 3.8 27b

More stories

  1. I built a desktop app where Claude Code designs machines, writes their firmware, and simulates both together — r/ClaudeAI
  2. Qwen 3.8 27B on one 5090: 22 GB VRAM, 175k context — a coding driver, not a Claude replacement — r/LocalLLM
  3. Qwen 3.8 with Claude is amazing — r/LocalLLM
  4. GPT-6 Sol and Luna now available on AI Gateway — Vercel Blog
  5. Sam Altman’s remarks at the United Nations Security Council — OpenAI News
  6. Anthropic launches Claude Opus 5.5 with stricter safeguards for cybersecurity — The Verge AI
  7. Alibaba unveils new AI chip to challenge NVIDIA, plans Qwen models with up to 10 trillion parameters — Mint AI
  8. Moonshot’s Kimi K3 lands on Amazon in key test for Chinese open-source AI revenue — South China Morning Post Tech

Get the daily brief of stories like this at 6:30 every morning →