AI Infra Summit: NVIDIA Vera Rubin and DSX Platform Advancements Showcase Energy Efficiencies of Optimizing Tokens Per Watt for AI Factories
AI Classified by Officially
Ian Buck, vice president of hyperscale and high-performance computing at NVIDIA, Tuesday spoke on AI factory efficiency at the AI Infra Summit, the Santa Clara Convention Center event that has morphed into a Coachella of infrastructure tech.
Before a packed audience — with more than 8,000 attendees this year, up from 3,500 last year — Buck discussed new collaborations across NVIDIA platforms and more.
- Amazon’s Annapurna Labs is working with NVIDIA on the NVHBM custom high-bandwidth memory technology.
- d-Matrix is integrating with NVLink Fusion to combine NVIDIA Vera CPUs with d-Matrix Raptor XPUs to deliver ultra low-latency inference at scale.
NVIDIA and partners unveiled new results as well:
- Emerald AI and NVIDIA demonstrated a commercial AI factory flexible-load program, working with Silicon Valley Power.
- Lambda improved performance per watt by 23% with NVIDIA DSX MaxLPS.
- Pinterest is using the NVIDIA Blackwell platform and NVIDIA Dynamo inference software to bring conversational AI to visual discovery.
The news comes as agentic AI is driving a new class of workloads that demand more performance, efficiency and scale from AI infrastructure.
NVIDIA addresses that challenge with a full-stack AI factory platform spanning Vera Rubin systems, Dynamo inference software, NeMo libraries and NVIDIA networking — including NVIDIA NVLink for scale-up computing, Spectrum-X Ethernet and ConnectX SuperNICs for connecting thousands of nodes, BlueField-powered context-memory storage and BlueField DPUs for infrastructure security.
“Infrastructure that’s fungible, that’s reliable, that’s going to last 10 years and really becomes an asset for computing the world’s computing problems in the world’s industries — they can build that with Vera Rubin, with DSX and all the innovations that we have here,” said Buck.
This is an extract. The publication continues at the source.
Read the original at the source: https://blogs.nvidia.com/blog/ai-infra-summit-vera-rubin-dsx-energy-efficiencies-tokens-per-watt-ai-factories/
Officially imported this from NVIDIA’s own source and shows an extract. If you work there, claiming the profile and verifying the domain lets you choose to show the full text here.
This publication has changed since it was first published
2 versions recorded. The original is kept in full — nothing is overwritten.
- v2 imported change on current
- v1 as first published on
Provenance
- Organization
- NVIDIA — imported from official source
- Official source
- https://blogs.nvidia.com/feed/ RSS
- Imported
- September 15, 2026 19:08
- Versions
- 2 recorded
- Identity
https://blogs.nvidia.com/?p=98178