AINewsnow

AI Infra Summit: NVIDIA Vera Rubin and DSX Platform Advancements Showcase Energy Efficiencies of Optimizing Tokens Per Watt for AI Factories

This story is from 2026-09-15. It is preserved in the archive; the latest stories are on the live feed.

Ian Buck, vice president of hyperscale and high-performance computing at NVIDIA, Tuesday spoke on AI factory efficiency at the AI Infra Summit, the Santa Clara Convention Center event that has morphed into a Coachella of infrastructure tech. Before a packed audience — with more than 8,000 attendees…

Read the full story at NVIDIA Blog ↗

Timeline · 4 reports

  1. 2026-09-16 15:26 · CoreWeave Blog
    CoreWeave Leads Cloud Providers in MLPerf® Inference v6.1 Performance with NVIDIA Blackwell Ultra
  2. 2026-09-16 15:00 · NVIDIA Blog
    NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut
  3. 2026-09-16 13:33 · CoreWeave Blog
    What It Takes to Bring Up a Multi-Rack NVIDIA Vera Rubin NVL72 Cluster
  4. 2026-09-15 16:55 · NVIDIA Blog
    AI Infra Summit: NVIDIA Vera Rubin and DSX Platform Advancements Showcase Energy Efficiencies of Optimizing Tokens Per Watt for AI Factories

More stories

  1. Amid growing AI fears, King Charles meets with industry leaders in Scotland — NPR Technology
  2. King Charles to press Nvidia, OpenAI, Anthropic leaders on AI safety at summit — CNBC Technology
  3. Huawei details AI accelerator roadmap, pulls in next-generation Ascend NPUs by several quarters — FP4 performance of the Ascend 960PR doubles expectations — Tom's Hardware
  4. Flyweight: open-source C++/CUDA engine for running MoE models bigger than your VRAM on one GPU + system RAM. First PyPI release, looking for contributors. — r/LocalLLaMA
  5. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM
  6. Huawei unveils latest tech to boost AI power in push to break China’s Nvidia reliance — South China Morning Post Tech
  7. Beginner confused about Ollama vs LM Studio vs llama.cpp vs vLLM vs Unsloth — can someone explain? — r/LocalLLM
  8. what's the state of the art recipe for running Qwen3.8-Flash-Next with a pair of 3090s and a ton of system RAM rn? — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →