newsfilter.io
Fireside Chat, Conference Presentation

Serge Palaric, NVIDIA & Antoine Vaissié, Google Cloud: Better together

  • Strategic Partnership Evolution: The collaboration between Google Cloud and NVIDIA has entered a new, more strategic phase focused on enterprise-scale AI, moving beyond simple hardware availability to deep engineering collaboration.
  • Blackwell GPU Availability:
    • Google Cloud is the first provider to offer NVIDIA's Blackwell architecture, specifically the B200 and GB200 GPUs.
    • A4 Instances: Powered by B200 GPUs, available for general AI workloads.
    • A46 Instances: Powered by GB200 GPUs, designed for high-performance scaling.
    • G4 Instances: Newly announced instances optimized for AI inference and advanced graphics.
    • RTX Pro 6000 Blackwell Server Edition: Available as a general-purpose GPU for enterprises requiring capabilities in graphics, AI inference, and digital twin simulations.
  • Industry Scaling Trends:
    • Shift in Workloads: The industry is transitioning from pre-training to "post-training," driven by the rise of reasoning models that generate significantly more tokens.
    • Agentic AI: The emergence of agentic AI involves multiple models interacting to generate vast amounts of data, increasing requirements for storage, management, and security.
  • NVIDIA Software & Services Integration:
    • NVIDIA AI Enterprise: Fully integrated into Google Cloud Marketplace, providing production-grade support, security, and management tools (e.g., moving from open-source development to secure production).
    • DGX Cloud: Available on Google Cloud, offering customers access to a supercomputer-like environment with pre-configured reference architectures, networking, and GPUs to accelerate initial deployment before migrating to standard cloud instances.
    • NVIDIA NIM (Inference Microservices): A suite of 150+ optimized containers for deploying models (including Google's Gemma) directly on GPUs, allowing developers to build applications without developing the underlying model infrastructure.
    • Licensing Model: NVIDIA AI Enterprise operates on a per-GPU subscription license, granting access to the software stack and dedicated support.
  • Data Sovereignty & Confidential Computing:
    • NVIDIA hardware and software are accessible via Google Distributed Cloud for local deployment requirements.
    • Specific sovereign initiatives include SOS (Sens), a joint venture between Google Cloud and Thales, offering advanced services locally in France.
  • Vertical-Specific Strategies:
    • Targeted Verticals: The partnership focuses on specific industries including Healthcare (genomics), Industrial, Finance, and Research.
    • Domain Expertise: NVIDIA and Google aim to understand industry-specific languages and workflows (e.g., "genetic" language in genomics) to provide relevant frameworks and use cases.
    • Enterprise Alignment: Successful AI adoption requires involving IT departments early to manage budgets and ensure operational integration, rather than relying solely on business or data science teams.
  • Forward-Looking Statements & Value Proposition:
    • Immediate Deployment: All mentioned Blackwell hardware, NVIDIA AI Enterprise, and NIMs are available for deployment "today," not in the future.
    • Risk Reduction: The integrated stack allows customers to deploy faster with reduced risk by utilizing pre-validated reference architectures and support models.
    • Engineering Collaboration: The relationship is built on joint engineering, testing, and qualification to solve complex, multi-threaded problems involving networking, storage, and compute.