NVIDIA’s Rubin Lands Inside Google’s Virtual Machine, Stretching Multi Site Clusters to Nearly 1 Million GPUs

NVIDIA’s Rubin Lands Inside Google’s Virtual Machine, Stretching Multi Site Clusters to Nearly 1 Million GPUs

By Ramish Zafar
Publication Date: 2026-04-27 19:09:00

Google and NVIDIA have teamed up to provide users with access to as much as one million NVIDIA GPUs to power up the freshly launched A5X instances. The announcement is part of the pair’s latest collaboration to reduce inference costs and improve token throughput. Their A5X system relies on NVIDIA’s network accelerators that enable the development of single and mutli-cluster computing infrastructure for AI workloads.

NVIDIA & Google to Allow Connecting Nearly a Million Rubin AI GPUs For A5X Instances Through Latest Collaboration

The A5X instances are Google’s latest products that are designed specifically to run agentic artificial intelligence workloads. They are part of Google’s AI Hypercomputer portfolio which also powers the firm’s Gemini platform and its consumer and enterprise AI offerings. As part of its latest announcements, Google announced a slew of upgrades to Hypercomputer which include new virtual machines powered by custom Arm-based CPUs, eight generation tensor processors, native PyTorch TPU support and the A5X instances.

These new capabilities are designed specifically to target agentic AI workloads which rely on a group of AI agents to focus on a piece-wise approach of solving a problem or a task. The A5X instances are the first ones from Google that are designed to work on NVIDIA’s latest Vera Rubin AI GPUs.

NVIDIA & Google’s AI Partnership Expand Virtual Machine GPU Use

According to the details, the A5X…