TLDR: Google Cloud is significantly advancing its artificial intelligence capabilities through expanded partnerships with Nvidia, focusing on the integration of Gemini models with Nvidia’s Blackwell GPUs and other hardware. This collaboration aims to enhance AI performance, security, and flexibility for both cloud and on-premises deployments, particularly benefiting enterprise customers in regulated sectors.
Google Cloud and Nvidia have announced a substantial expansion of their strategic partnership, ushering in a new era of advanced artificial intelligence solutions. This deepened alliance centers on optimizing Google’s powerful Gemini and Gemma AI models for Nvidia’s cutting-edge GPU infrastructure, including the new Blackwell GPUs, and extending their deployment capabilities across various environments.
One of the most significant developments is the ability to deploy Google’s Gemini models on-premises using Nvidia’s Blackwell GPUs via the Google Distributed Cloud. This innovation grants enterprises unprecedented control over their data and AI models, addressing critical compliance and privacy requirements in highly regulated sectors such as healthcare, defense, and finance. Uttara Kumar, senior product marketing manager for Nvidia, highlighted that this move ‘unlocks agentic AI for customers’ within their own secure data centers.
Performance optimization is a key focus of this collaboration. Gemini inference workloads are now designed to run with greater speed and efficiency across Google Cloud’s Vertex AI and the Google Distributed Cloud, leveraging Nvidia’s hardware acceleration. Furthermore, the Gemma family of lightweight, open models has been meticulously tuned using the Nvidia TensorRT-LLM library. These optimized Gemma models will be available as easy-to-deploy Nvidia NIM microservices, simplifying AI deployment and maximizing resource utilization for developers.
Beyond on-premises deployment and performance enhancements, the partnership also extends to confidential computing with Nvidia H100 GPUs, reinforcing Google Cloud’s commitment to robust security for AI workloads. The broader scope of this alliance, initially announced in March 2024, includes the adoption of the new Nvidia Grace Blackwell AI computing platform and the Nvidia DGX Cloud service on Google Cloud, which is now generally available.
Google Cloud CEO Thomas Kurian emphasized the comprehensive nature of this partnership, stating, ‘The strength of our long-lasting partnership with NVIDIA begins at the hardware level and extends across our portfolio – from state-of-the-art GPU accelerators, to the software ecosystem, to our managed Vertex AI platform.’ Nvidia’s founder and CEO, Jensen Huang, echoed this sentiment, noting that enterprises are seeking solutions to rapidly leverage generative AI, and this partnership provides an ‘open, flexible platform to easily scale generative AI applications.’
Also Read:
- Nvidia Invests $5 Billion in Intel, Forging Strategic AI and PC Partnership
- Cloud Platforms Emerge as the Indispensable Foundation for AI Data Management
This integrated approach, combining Nvidia’s advanced hardware with Google Cloud’s software and AI models, provides a flexible, scalable, and secure foundation for businesses to accelerate their AI projects, whether in the cloud or on-premises. The collaboration is poised to meet the growing demand for high-performance AI infrastructure, offering solutions that deliver control, performance, and trust to enterprise-grade AI systems.


