Old datacenter GPUs aren't dead: PXA v3 runs 27B at 97 t/s on two V100s, 86 t/s on ONE, Gemma 4 at 169 t/s, and a 124B-class MoE on a single P100 + RAM (free, open engine + one-click GUI)
Hey r/LocalLLaMA 👋 We're PXA Network . We build PXA , an inference engine tuned for Pascal and Volta GPUs (P100, V100, P40, 1080 Ti), the $100–300 eBay cards everyone calls too old. PXA v3 just shipped, and everything runs through PXA Control , our one-click web app. ⚡ Speed (PXA v3, measured on o…
Read the full story at r/LocalLLM ↗