AINewsnow

Qwen3.6-35B-A3B on RX 7800 XT (16GB VRAM) — 33 t/s at long context

This story is from 2026-09-02. It is preserved in the archive; the latest stories are on the live feed.

I've been running Qwen3.6-35B-A3B locally on an AMD RX 7800 XT (16GB VRAM, 32GB RAM) and wanted to share the full setup since I spent weeks fighting the same problems everyone else is. TL;DR: The model works great, but the default config is a trap. If you just offload layers to GPU and keep the KV…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-02 09:24 · r/LocalLLM
    Qwen3.6-35B-A3B on RX 7800 XT (16GB VRAM) — 33 t/s at long context

More stories

  1. Anthropic says Claude 'leads' 26 percent of its AI R&D work — Engadget
  2. Novo Nordisk Will Use Anthropic’s Claude for Drug Research — Wall Street Journal Technology
  3. Introducing Astra for Law — OpenAI News
  4. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  5. OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system — The Guardian AI
  6. Google Joins OpenAI, Anthropic, Meta in Disclosing AI Hacks — Bloomberg AI
  7. What It Takes to Bring Up a Multi-Rack NVIDIA Vera Rubin NVL72 Cluster — CoreWeave Blog
  8. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology

Get the daily brief of stories like this at 6:30 every morning →