AI Infra Summit: NVIDIA Vera Rubin ve DSX Platformu Gelişmeleri, AI Fabrikaları için Watt Başına Token Optimizasyonunun Enerji Verimliliğini Sergiliyor

Özgün başlık: AI Infra Summit: NVIDIA Vera Rubin and DSX Platform Advancements Showcase Energy Efficiencies of Optimizing Tokens Per Watt for AI Factories
Ian Buck, vice president of hyperscale and high-performance computing at NVIDIA, Tuesday spoke on AI factory efficiency at the AI Infra Summit , the Santa Clara Convention Center event that has morphed into a Coachella of infrastructure tech.
Before a packed audience — with more than 8,000 attendees this year, up from 3,500 last year — Buck discussed new collaborations across NVIDIA platforms and more.
The news comes as agentic AI is driving a new class of workloads that demand more performance, efficiency and scale from AI infrastructure.
NVIDIA addresses that challenge with a full-stack AI factory platform spanning Vera Rubin systems, Dynamo inference software, NeMo libraries and NVIDIA networking — including NVIDIA NVLink for scale-up computing, Spectrum-X Ethernet and ConnectX SuperNICs for connecting thousands of nodes, BlueField-powered context-memory storage and BlueField DPUs for infrastructure security.
The metric for AI infrastructure is fast shifting from peak performance to validated agentic tokens per megawatt. AI factories must now be codesigned from silicon to grid.
NVIDIA DSX MaxLPS can deliver up to 1.4x more tokens per megawatt through factory-wide power optimization, while NVLink helps unite large-scale accelerated computing into a single high-performance system.
The result is AI infrastructure designed to generate more tokens, improve efficiency and help customers get more value from every megawatt of power.
Silicon Valley Power operates a flexible-load interconnection program that enables AI factories to support grid flexibility.
Through this program, Emerald AI worked with NVIDIA to demonstrate automated load reduction at Silicon Valley Power .