Follow-up on my 3x Tesla P100 budget build: cooling, power draw, PCIe lanes and tokens/sec!
Update: I found a mistake in my three-card numbers. Please disregard these parts of the post below: The 3x P100 row in the tokens/sec table (I had ~21 t/s generation and ~67 t/s prompt processing) The "three cards was about 4x slower than two" bullet, and my guess about why The three-card power (~1…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-24 03:53 · r/LocalLLM
Follow-up on my 3x Tesla P100 budget build: cooling, power draw, PCIe lanes and tokens/sec!