AWS will deploy 2 million more NVIDIA GPUs across its global infrastructure in 2027 and 2028.
The expansion follows its earlier plan to add more than 1 million NVIDIA GPUs starting in 2026.
The new capacity will include Blackwell Ultra, Rubin and Rubin Ultra GPUs for agentic AI, scientific discovery, enterprise automation and physical AI.
The companies will also expand their work across processors, networking, AI models and data processing.
AWS plans to introduce NVIDIA Vera processor infrastructure and explore using NVIDIA’s high-bandwidth memory with future Trainium chips.
They also plan to build AI factories for the US government, including 100,000 GPUs for federal and national security workloads.
“Customers want the freedom to choose the best tools for their AI workloads, and they want confidence that everything works seamlessly together.
That’s why we’ve invested deeply with NVIDIA to make AWS the best place to run NVIDIA AI technologies, optimising performance across our infrastructure from networking and security to deployment.”
“NVIDIA and AWS have built one of the great growth engines of the AI era, and demand is running ahead of every forecast. For 16 years, we have scaled NVIDIA computing in the cloud together.
Now, we are expanding our partnership across the full stack — GPUs, CPUs, networking, open models and software — to make agentic and physical AI real at an unprecedented pace and scale that only AWS and NVIDIA can deliver.”
NVIDIA’s Nemotron open models will remain available through Amazon Bedrock and SageMaker.
The companies are also bringing NVIDIA technology to Amazon EMR and OpenSearch to speed up data processing and vector indexing.
Source link







