Amazon Web Services (AWS) and NVIDIA announced a major expansion of their strategic collaboration. Building on 16 years of joint innovation, the companies plan to deploy 2 million additional NVIDIA GPUs across AWS’s global infrastructure and deepen their work together across AI factories, CPUs, networking, open models, data processing, and robotics—delivering co-engineered solutions that help customers accelerate AI development at unprecedented scale.
The expansion complements Amazon’s own custom silicon, giving customers the freedom to choose the best compute for their specific workloads—whether that’s NVIDIA GPUs, AWS Trainium chips, or both working together.
AWS and NVIDIA have worked together to bring cutting-edge AI capabilities to customers around the world for nearly two decades. In fact, the two companies together launched the world’s first GPU-accelerated cloud instance on AWS, and today AWS offers the widest range of NVIDIA GPU solutions for customers.
Now, as demand for AI accelerates, the two companies are taking that collaboration to a new level. AI workloads are scaling at a rapid pace—from how models are trained and deployed, to how data is processed and used to power intelligent applications. Customers are moving from pilot to production across agentic AI, scientific discovery, enterprise automation, and robotics, and they need infrastructure that can keep pace.
“Customers want the freedom to choose the best tools for their AI workloads, and they want confidence that everything works seamlessly together,” said Matt Garman, CEO of AWS. “That’s why we’ve invested deeply with NVIDIA to make AWS the best place to run NVIDIA AI technologies, optimizing performance across our infrastructure from networking and security to deployment. This expanded collaboration gives frontier labs, enterprises, and governments even more ways to build and deploy AI on AWS.”
“NVIDIA and AWS have built one of the great growth engines of the AI era, and demand is running ahead of every forecast,” said Jensen Huang, founder and CEO of NVIDIA. “For 16 years, we have scaled NVIDIA computing in the cloud together. Now we are expanding our partnership across the full stack—GPUs, CPUs, networking, open models and software—to make agentic and physical AI real at an unprecedented pace and scale that only AWS and NVIDIA can deliver. This expansion reflects customers’ demand for NVIDIA’s platform on AWS.”

