Upgrade from 32 -> 48 -> 64gb
Currently running Qwen 3.8 27b via Ninfer on 5090 with 240k context. I have one more 8x and one 4x PCI slots. Thinking about adding one or two more GPUs. But what would be actual level up? Would the models that require 48 or 64gb are noticeably performing better? Mostly using for coding and some pe…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-01 22:16 · r/LocalLLM
Upgrade from 32 -> 48 -> 64gb
More stories
- Strands Labs, AWS's experimental agent-development project, unveils Strands Decider 2B, a free, open-source Jev competitor fine-tuned from an Alibaba Qwen base (Carl Franzen/VentureBeat) — Techmeme
- Two open-weights releases: Victoria (Qwen3.8-Flash-Next with 44% of experts cut, 70% Terminal-Bench 2.1, GGUF included) and Maple (a Canada-first fine-tune) — r/LocalLLM
- Qwen flash next on 12+16gb vram, and 32gb ram viable? — r/LocalLLM
- Train Edit Loras for Qwen image 2.1 in Fizgig 6.6.0 — r/StableDiffusion
- add GLM-5.3-Flash (GLM5-Next) support by timkhronos · Pull Request #27773 · ggml-org/llama.cpp — r/LocalLLaMA
- Browser FPS with 3D models, textures and SFX generated locally on one GPU, plus a local Qwen 27B for part of the code: my pipeline and what failed — r/LocalLLM
- Continuity update: screen replacement, image to 3D, Qwen Image 2.1, and a lot more since 3.0 — r/StableDiffusion
- Is anyone else running insanely long unattended loops? — r/AI_Agents
Get the daily brief of stories like this at 6:30 every morning →