AINewsnow

Follow-up on my 3x Tesla P100 budget build: cooling, power draw, PCIe lanes and tokens/sec!

Update: I found a mistake in my three-card numbers. Please disregard these parts of the post below: The 3x P100 row in the tokens/sec table (I had ~21 t/s generation and ~67 t/s prompt processing) The "three cards was about 4x slower than two" bullet, and my guess about why The three-card power (~1…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-24 03:53 · r/LocalLLM
    Follow-up on my 3x Tesla P100 budget build: cooling, power draw, PCIe lanes and tokens/sec!

More stories

  1. If tomorrow AGI / ASI is able to develop a device you plug into your car that let's your car be operated and driven by AGI, would you trust it enough to let it drive you around? — r/agi
  2. looking for recommendation for dual gpu setup, with lga 1700 — r/LocalLLM
  3. 2× Tesla P100 (2016 cards) in 2026: 110 tok/s on a 30B MoE, 16 tok/s at 1M context — r/LocalLLM
  4. P100 local llm budget build. — r/LocalLLM
  5. GPT-6 Sol and Luna now available on AI Gateway — Vercel Blog
  6. Bringing Private Processing to Meta AI Glasses — Engineering at Meta
  7. Gemini 3.8 text-to-speech models now available on AI Gateway — Vercel Blog
  8. Sam Altman’s remarks at the United Nations Security Council — OpenAI News

Get the daily brief of stories like this at 6:30 every morning →