AINewsnow

2 GPUs, but one in PCIe 3.0 1x slot

I have a 'gaming' motherboard which provides one PCIe 5.0 x16. The two remaining slots are full-length x16 but physically wired for only x1, at gen 3.0 speed - about 1 GB/s... I'm running llama.cpp and the two GPUs are AMD 9070XT (16GB VRAM) and a R9700 (32GB VRAM) I just ordered. Due to the extrem…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-27 13:04 · r/LocalLLM
    2 GPUs, but one in PCIe 3.0 1x slot

More stories

  1. Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Vision-Language Models With Up to 3.13x Faster Decoding — MarkTechPost
  2. Ternary Bonsai 2 27B at up to 532 tok/s on one RTX 4090, native Windows: MTP + n-gram speculative decoding in a from-scratch CUDA engine — r/LocalLLM
  3. Updated from 3x3090(2x3090, 1x3090TI) to 2x5090 — r/LocalLLaMA
  4. is switching from llama cpp to vllm worth it — r/LocalLLaMA
  5. Adding logit penalty for "wait", "maybe" and "perhaps" to Qwen models improves their accuracy — r/LocalLLaMA
  6. I ran the Laya model on llama.cpp to test the game of Snake, and the results were excellent. — r/LocalLLM
  7. Qwen3.8-Flash-Next on 12GB VRAM - 65 tokens per second — r/LocalLLM
  8. JiRackUltra_1b Runs AI Routing on Any Laptop Without a GPU — AlphaSignal

Get the daily brief of stories like this at 6:30 every morning →