Nvidia announced its next-generation AI platform, Vera Rubin, during the CES 2026 keynote, introducing a rack-scale architecture designed to dramatically improve AI inference and training performance for data centers. The Rubin platform, which includes the Rubin GPU, Vera CPU, NVLink 6 switch, ConnectX-9 SuperNIC, BlueField-4 DPU, and Spectrum-6 Ethernet switch, promises up to 50 PFLOPS of inference performance per GPU and 288GB of HBM4 memory, delivering up to 5x greater inference performance and 10x lower cost per token compared to the previous Blackwell generation. Nvidia's CEO Jensen Huang emphasized the importance of "extreme co-design" across hardware and software to eliminate bottlenecks and reduce operational costs, positioning Rubin as a key enabler for scaling AI workloads across industries.
The Vera Rubin NVL72 rack is set to address the growing demand for AI compute in data centers, supporting advanced applications in robotics, autonomous vehicles, and large language models. Nvidia also highlighted its expanding open model portfolio for sectors such as healthcare, climate science, and robotics, and introduced the Alpamayo VLA model for autonomous driving, which will debut in the new Mercedes-Benz CLA. The Rubin platform is expected to be available in the second half of 2026, marking a significant leap in AI infrastructure capabilities and economic efficiency for enterprise and hyperscale deployments.

Track how attackers are adapting to this technology.
4 events from the most recent confirmed update back to the earliest known activity.
Nvidia stated that the Vera Rubin NVL72 platform is expected to reach volume production in the second half of 2026. The production timeline accompanied the CES 2026 launch announcement and framed the platform's commercial availability.
Alongside the Rubin platform announcement, Nvidia introduced the Alpamayo vision-language-action model for autonomous vehicles, designed to interpret sensor data and make driving decisions. Nvidia said the new Mercedes-Benz CLA would be the first car to feature Alpamayo in the US later in 2026.
Nvidia said the Vera Rubin NVL72 rack would deliver up to 3.6 exaFLOPS of inference and 2.5 exaFLOPS of training performance per rack, while cutting AI token generation costs by up to 10x versus the prior Blackwell generation. The company also highlighted new memory and networking features, expanded trusted execution environments, and serviceability improvements such as modular design and zero-downtime maintenance.
At CES 2026, Nvidia CEO Jensen Huang introduced the next-generation Vera Rubin/Rubin AI data center platform, featuring the Vera CPU, Rubin GPU, NVLink 6, ConnectX-9 SuperNIC, BlueField-4 DPU, and Spectrum networking components. Nvidia positioned it as a rack-scale, tightly integrated AI infrastructure platform for training and inference workloads.
Vulnerabilities, threat actors, malware, products, organizations, and breaches Mallory has linked to this story.
Follow how adversaries are adapting to this technology, and where it touches your stack today.
2 references tracked. Mallory keeps watching after this page renders.
cio.com
Open sourcetomshardware.com
Open sourceMap indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.