LIVE
Google DeepMind Introduces Gemini 3.8 Live and Its Extended Thinking Variant15/09/26 · Google DeepMind|What's at stake in AI's trillion-dollar infrastructure bet15/09/26|A Flaw in Chain-of-Thought Safety Monitoring14/09/26|Stellar Colosseum: A Multi-Agent System for Long-Horizon Mathematical Research14/09/26|Apple Code Hints Siri Could Be Swapped for ChatGPT or Claude14/09/26 · Apple|Anthropic says Houthi-linked actors used Claude Code for missile guidance software13/09/26 · Anthropic|Yoshua Bengio examines why AI agents lie, cheat and coordinate13/09/26|Terence Tao warns of a 'severe misalignment' in OpenAI's mathematical claims11/09/26 · OpenAI|Terence Tao warns of a 'severe misalignment' of AI in mathematics11/09/26|OpenAI's Navier-Stokes release comes with a Lean 4 formal proof10/09/26 · OpenAI|OpenAI publishes documentation for its Agents API10/09/26 · OpenAI|OpenAI Launches the Agents API, a Managed Service for Cloud Agents10/09/26 · OpenAI|Google DeepMind Introduces Gemini 3.8 Live and Its Extended Thinking Variant15/09/26 · Google DeepMind|What's at stake in AI's trillion-dollar infrastructure bet15/09/26|A Flaw in Chain-of-Thought Safety Monitoring14/09/26|Stellar Colosseum: A Multi-Agent System for Long-Horizon Mathematical Research14/09/26|Apple Code Hints Siri Could Be Swapped for ChatGPT or Claude14/09/26 · Apple|Anthropic says Houthi-linked actors used Claude Code for missile guidance software13/09/26 · Anthropic|Yoshua Bengio examines why AI agents lie, cheat and coordinate13/09/26|Terence Tao warns of a 'severe misalignment' in OpenAI's mathematical claims11/09/26 · OpenAI|Terence Tao warns of a 'severe misalignment' of AI in mathematics11/09/26|OpenAI's Navier-Stokes release comes with a Lean 4 formal proof10/09/26 · OpenAI|OpenAI publishes documentation for its Agents API10/09/26 · OpenAI|OpenAI Launches the Agents API, a Managed Service for Cloud Agents10/09/26 · OpenAI|
ReleaseNVIDIA

NVIDIA Details Energy Efficiency Gains of Vera Rubin and DSX Platform

At the AI Infra Summit, NVIDIA showcased progress on its Vera Rubin and DSX platforms, emphasizing tokens generated per watt as a key new metric for AI data centers.

September 15, 20263 min readPublished byNVIDIA

Ian Buck, NVIDIA's vice president of hyperscale and high-performance computing, spoke Tuesday at the AI Infra Summit, an event held at the Santa Clara Convention Center whose attendance has more than doubled year over year, growing from roughly 3,500 to over 8,000 participants. That surge reflects the industry's intensifying focus on infrastructure built to support generative AI workloads at scale.

A central theme of his remarks was the concept of the AI factory, a framing NVIDIA has been promoting for several quarters to describe data centers optimized specifically for token production rather than general-purpose computing. Buck detailed progress on the Vera Rubin and DSX platforms, positioned as successors to the current Blackwell architecture, highlighting improvements in tokens generated per watt of energy consumed — a metric the company is increasingly pushing as a standard benchmark for comparing AI infrastructure efficiency.

This emphasis on energy efficiency comes as power consumption at AI-focused data centers has become a growing concern for operators, utilities, and regulators alike. By framing efficiency gains in terms of watts rather than raw compute output, NVIDIA is signaling that upcoming chip generations aim to deliver more useful output without a proportional rise in energy costs for operators running these facilities.

While precise performance figures for Vera Rubin remain only partially disclosed, the presentation reinforces NVIDIA's broader strategy: positioning its platforms not merely as compute accelerators, but as components of infrastructure designed end-to-end for the industrial scale of generative AI deployment.

Tags
nvidiadata-centersenergy-efficiencyvera-rubinai-infrastructuregpu

Read also

NVIDIA Details Energy Efficiency Gains of Vera Rubin and DSX Platform · nAIvigate