AWS, NVIDIA Expand AI Infrastructure Partnership with Additional 2 Million GPUs
AWS and NVIDIA are expanding their partnership to meet growing demand for AI infrastructure. The companies plan to deploy an additional 2 million NVIDIA GPUs across AWS's global infrastructure by 2027-2028, building on previous collaborations that have already accelerated customer adoption of NVIDIA-accelerated compute on AWS.
AI workloads are scaling rapidly, and customers need broader model choice, faster data pipelines, and new capabilities for emerging use cases like physical AI. To meet this demand, AWS and NVIDIA will deepen their work together across AI factories, CPUs, networking, open models, data processing, and robotics, delivering co-engineered AI solutions that enable customers to accelerate AI development and deployment at unprecedented scale.
The expanded collaboration includes plans to bring NVIDIA Vera CPU-based infrastructure to AWS, integrate the NVIDIA platform with the AWS Nitro System and Elastic Fabric Adapter (EFA) for enhanced security and reliability, and continue to support NVIDIA Nemotron open models on Amazon Bedrock and Amazon SageMaker. The companies will also accelerate data processing and vector indexing on Amazon EMR and Amazon OpenSearch with NVIDIA cuDF and cuVS CUDA-X libraries for faster, more cost-efficient analytics and AI applications.