CoreWeave Delivers Breakthrough AI Performance with NVIDIA GB200 and H200 GPUs in MLPerf Inference v5.0
This story is from 2026-09-08. It is preserved in the archive; the latest stories are on the live feed.
CoreWeave achieved top MLPerf v5.0 AI inference results, delivering 800 TPS on Llama 3.1 405B with NVIDIA GB200 and 33,000 TPS on Llama 2 70B with H200 GPUs, marking significant performance gains.
Read the full story at CoreWeave Blog ↗
Timeline · 1 report
- 2026-09-08 13:59 · CoreWeave Blog
CoreWeave Delivers Breakthrough AI Performance with NVIDIA GB200 and H200 GPUs in MLPerf Inference v5.0