AINewsnow

Qwen3.8 27B Q8_0 @ W7900 - performance question?

Are these reasonable numbers? Qwen3.8 27B Q8_0 Radeon PRO W7900 48GB / ROCm-HIP (gfx1100) FP16 KV / custom tiled GPU attention / all layers GPU DFlash2 Q4_K_M / adaptive drafting, up to 7 tokens / shared head 262,144-token context / 1 concurrent session Razer Core X Chroma external GPU enclosure --…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-10 21:37 · r/LocalLLM
    Qwen3.8 27B Q8_0 @ W7900 - performance question?

More stories

  1. Introducing GPT-6 in ChatGPT with Intelligent UI — OpenAI YouTube
  2. An Anthropic AI model sent a false homicide tip to Philadelphia police — TechCrunch AI
  3. Anthropic bans users from ‘needless abusive or cruel behavior’ towards Claude — The Guardian AI
  4. Philadelphia police receive false homicide tip from Anthropic AI model — The Hill Technology
  5. Impactful scheduling for GPU clusters — Allen Institute for AI (Ai2)
  6. Sophos cuts threat investigation time by 96% with OpenAI Daybreak — OpenAI News
  7. Welcome to Gemini at Work 2026: Introducing the Gemini agent — Google Cloud AI Blog
  8. Qwen Image 2.1 Turbo Released -- Hugging Face — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →