Key insights
- NVIDIA and Google Cloud are expanding their AI collaboration, introducing new infrastructure and services. The new systems offer significant improvements in inference cost and throughput. OpenAI and Thinking Machines Lab are already utilizing these systems. Google Gemini models are now in preview on Google Distributed Cloud. This collaboration signals continued growth and innovation in the AI sector, positively influencing market sentiment towards AI-related stocks.

Investing.com -- NVIDIA and Google Cloud announced on Wednesday an expansion of their partnership to advance agentic and physical AI capabilities, introducing new infrastructure and services at Google Cloud Next in Las Vegas.
The companies unveiled NVIDIA Vera Rubin-powered A5X bare-metal instances, which can scale up to 960,000 NVIDIA Rubin GPUs in a multisite cluster. The A5X instances use NVIDIA ConnectX-9 SuperNICs combined with Google Virgo networking, scaling to up to 80,000 NVIDIA Rubin GPUs within a single site cluster.
The new systems deliver up to 10 times lower inference cost per token and 10 times higher token throughput per megawatt compared to the prior generation, according to the announcement.
Google Cloud’s NVIDIA Blackwell portfolio includes A4 VMs with NVIDIA HGX B200 systems, rack-scale A4X VMs with NVIDIA GB200 NVL72 and A4X Max NVIDIA GB300 NVL72 systems, and fractional G4 VMs with NVIDIA RTX PRO 6000 Blackwell Server Edition GPUs.
OpenAI is running large-scale inference on NVIDIA GB300 and GB200 NVL72 systems on Google Cloud for some of its inference workloads, including for ChatGPT. Thinking Machines Lab is scaling its Tinker API on A4X Max VMs with GB300 NVL72 systems.
Google Gemini models running on NVIDIA Blackwell and Blackwell Ultra GPUs are now in preview on Google Distributed Cloud. The companies also introduced Confidential G4 VMs with NVIDIA RTX PRO 6000 Blackwell GPUs, marking the first confidential computing offering of NVIDIA Blackwell GPUs in the cloud.
NVIDIA Nemotron 3 Super is now available on Gemini Enterprise Agent Platform. Google Cloud and NVIDIA introduced a new managed reinforcement learning API built with NVIDIA NeMo RL for accelerating training at scale.
CrowdStrike uses NVIDIA NeMo open libraries to generate synthetic data and fine-tune Nemotron and other open large language models for cybersecurity applications, running on Managed Training Clusters on Gemini Enterprise Agent Platform with NVIDIA Blackwell GPUs.
Solutions from Cadence and Siemens Digital Industries Software are now available on Google Cloud, accelerated on NVIDIA AI infrastructure. NVIDIA Omniverse libraries and the NVIDIA Isaac Sim robotics simulation framework are available on Google Cloud Marketplace.
NVIDIA received Google Cloud Partner of the Year recognition in two categories: AI Global Technology Partner and Infra Modernization Compute.
This article was generated with the support of AI and reviewed by an editor. For more information see our T&C.