proomt

Search

Search posts, papers, and topics

All posts

Nvidia7 min readintermediate

AI Infra Summit: NVIDIA Vera Rubin and DSX Platform Advancements Showcase Energy Efficiencies of Optimizing Tokens Per Watt for AI Factories

Summary

NVIDIA showcased new power‑aware AI‑factory features at the AI Infra Summit. DSX MaxLPS and Flex improve token‑per‑watt efficiency by up to 40%, while the Vera Rubin NVL72 platform delivers up to 30× higher throughput per megawatt on agentic workloads.

  • DSX MaxLPS lets 19 nodes run in the power budget of 16 full‑power nodes, boosting token throughput 24% and performance per watt 23%.
  • Vera Rubin NVL72 with DSX MaxLPS can fit up to 40% more GPUs in the same site‑power envelope, delivering up to 35% higher token throughput.
  • SemiAnalysis AgentX benchmark shows Vera Rubin NVL72 achieving up to 30× higher throughput per megawatt and 45× lower cost per million tokens for agentic workloads.
  • DSX Flex enables AI factories to automatically throttle low‑priority jobs in response to grid demand‑response signals, turning data centers into flexible load resources.

AI infrastructure engineers and data‑center operators should care because these efficiency gains directly reduce energy costs and enable larger LLM deployments within existing power limits.

6/10

Related reading

  1. From Megawatts to Tokens: How NVIDIA Maximizes AI Factory Production

    NVIDIA’s DSX platform lets AI data‑centers shift workloads in response to grid signals, squeezing ~24% more token throughput (4 M→5 M tps) and ~23% better performance‑per‑watt on a fixed megawatt budget. The first production demo used Emerald AI’s Conductor to drop a 4 MW load to 3 MW in under a minute without interrupting high‑priority jobs. DSX MaxLPS reallocates headroom across HGX B200 server…

    Nvidianvidia.com5 min
  2. NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut

    NVIDIA’s Vera Rubin NVL72 AI inference system shows up to 3.7× higher throughput than the prior GB300 NVL72 on MLPerf v6.1 benchmarks (Qwen3‑VL, DeepSeek‑R1), achieves 99% scaling efficiency across 288 GPUs, and benefits from software optimizations (NVFP4 precision, kernel fusion, disaggregated serving). The post is a product announcement with concrete benchmark numbers but limited technical dept…

    Nvidianvidia.com4 min
  3. 5 Companies Using NVIDIA AI for Clean Energy

    Nvidia’s blog spotlights five companies that are using Nvidia AI platforms to accelerate clean‑energy projects—from grid interconnection and nuclear plant operations to off‑grid AI data‑center power, advanced reactors, and fusion tokamaks. The article is a marketing summary and provides few technical details.

    Nvidianvidia.com4 min