Amazon Web Services (AWS), an Amazon.com, Inc. company (NASDAQ:AMZN), and NVIDIA (NASDAQ:NVDA) today announced a major expansion of their strategic collaboration to meet surging global demand for AI infrastructure as demand continues to accelerate. Building on already-rapid customer adoption of NVIDIA-accelerated compute on AWS, the companies plan to deploy 2 million additional NVIDIA GPUs across AWS’s global infrastructure and deepen their work together across AI factories, CPUs, networking, open models, data processing and robotics, delivering co-engineered AI solutions that enable customers to accelerate AI development and deployment at unprecedented scale.

AI workloads are scaling at a swift pace, from how models are trained and run, to how data is processed, indexed and used to power intelligent applications. Customers are moving from pilot to production and scaling workloads across agentic AI, scientific discovery, enterprise automation and robotics. They need broader model choice, faster data pipelines and new capabilities for emerging use cases like physical AI. They also need confidence that the underlying infrastructure can keep pace with their own ability to innovate while maintaining the highest level of security and reliability for mission-critical workloads.

To meet this surging demand from frontier labs, global enterprises, startups and governments, AWS and NVIDIA are building on 16 years of joint innovation to expand AI compute capacity and bring new co-engineered solutions to customers faster. As part of the expanded collaboration, the companies are working to:

  • Deploy 2 million additional NVIDIA GPUs across AWS’s global infrastructure in 2027-2028
  • Bring NVIDIA Vera CPU‑based infrastructure to AWS
  • Extend NVIDIA NVLink Fusion™ with custom NVIDIA high‑bandwidth memory (NVHBM)
  • Build AI factories for the U.S. government, including 100,000 GPUs on secure AWS infrastructure for running federal and national‑security workloads
  • Integrate the NVIDIA platform with the AWS Nitro System and Elastic Fabric Adapter (EFA) for enhanced security and reliability
  • Continue to support NVIDIA Nemotron™ open models on Amazon Bedrock and Amazon SageMaker, giving customers more open model choice
  • Accelerate data processing and vector indexing on Amazon EMR and Amazon OpenSearch with NVIDIA cuDF and cuVS CUDA-X™ libraries for faster, more cost‑efficient analytics and AI applications
  • Further advance robotics workloads through Amazon Robotics’ adoption of NVIDIA’s physical AI platform, speeding innovation in warehouse automation and next‑generation robots