Amazon Web Services (AWS) has signed on to expand its fleet of Nvidia GPUs.
The company has agreed to deploy an additional 2 million GPUs across its global infrastructure, and the two companies will also deepen their work on AI factories, CPUs, networking, open models, data processing, and robotics.
AWS will deploy the GPUs through 2027 and 2028. In addition, the cloud giant will bring Nvidia Vera CPU-based infrastructure to AWS, extend Nvidia NVLink Fusion with custom Nvidia high-bandwidth memory, and build AI factories for the US government with 100,000 GPUs on secure AWS infrastructure.
Nvidia's Nemotron models will also be offered via AWS, and the two will further robotics efforts with Amazon Robotics to adopt Nvidia's physical AI platform.
During Nvidia GTC 2026, AWS said it would bring online more than one million Nvidia GPUs to its platform in 2026, but demand has since outpaced this. The two million additional GPUs will include Blackwell Ultra, Rubin, and Rubin Ultra GPUs.
The company will also expand Blackwell capacity with the Nvidia RTX Pro 4500 Blackwell Server Edition GPUs for its EC2 G7 instances, which offer 4.6x AI inference performance and 2.1x graphics performance compared to previous-generation G6 instances
“Customers want the freedom to choose the best tools for their AI workloads, and they want confidence that everything works seamlessly together,” said Matt Garman, CEO of AWS. “That’s why we’ve invested deeply with Nvidia to make AWS the best place to run Nvidia AI technologies, optimizing performance across our infrastructure from networking and security to deployment. This expanded collaboration gives frontier labs, enterprises and governments even more ways to build and deploy AI on AWS.”
“Nvidia and AWS have built one of the great growth engines of the AI era, and demand is running ahead of every forecast,” said Jensen Huang, founder and CEO of Nvidia. “For 16 years, we have scaled Nvidia computing in the cloud together. Now, we are expanding our partnership across the full stack — GPUs, CPUs, networking, open models and software — to make agentic and physical AI real at an unprecedented pace and scale that only AWS and Nvidia can deliver. This expansion reflects customers’ demand for Nvidia's platform on AWS.”
While AWS continues to expand its fleet, the company's CEO Matt Garman has previously said that the company is still running the six-year-old A100s, having "never retired an A100 server."
Nvidia published its earnings results for Q2 2026 this week. The company posted $96.2 billion in revenue for Q2, up 18 percent from the previous quarter and up 106 percent from a year ago, while GAAP and non-GAAP gross margins were both 75 percent.
Comments