Skip to main content
Back to News
NVIDIA and AWS deepen their alliance with 2 million more GPUs, Vera CPU infrastructure, and a government AI factory.
Partnership
2 min read
US

NVIDIA and AWS deepen their alliance with 2 million more GPUs, Vera CPU infrastructure, and a government AI factory.

The AMW Read

Extends the known AWS-NVIDIA compute buildout with an explicit 2M-GPU hyperscaler-scale commitment, new Vera CPU/NVLink Fusion/NVHBM silicon integration, and a dedicated federal AI factory spanning infrastructure, data, and robotics workloads.
NoveltySignificance
AI Infra · Player MapCompute EconomicsSilicon Substrate

NVIDIA and AWS deepen their alliance with 2 million more GPUs, Vera CPU infrastructure, and a government AI factory.

AWS and NVIDIA announced an expansion of their 16-year partnership, committing 2 million additional NVIDIA GPUs — including Blackwell Ultra, Rubin, and Rubin Ultra — across AWS's global infrastructure between 2027 and 2028, on top of the 1 million-plus GPUs pledged at GTC 2026. The expansion brings NVIDIA Vera CPU-based infrastructure to AWS, extends NVLink Fusion with NVIDIA's custom high-bandwidth memory (NVHBM) for AWS's next-generation Trainium chips, and builds a secure AI factory with 100,000 GPUs for U.S. federal and national-security workloads rated for Impact Level 6 classification. AWS will be the first major cloud provider to offer NVIDIA's RTX PRO 4500 Blackwell Server Edition GPUs via new EC2 G7 instances, and will keep supporting NVIDIA's open Nemotron models on Bedrock and SageMaker alongside GPU-accelerated data processing on EMR and OpenSearch.

The scale signals that AI compute demand is still outrunning hyperscaler capacity plans announced only months earlier — AWS says demand has already exceeded what it projected at GTC 2026. Weaving Vera CPUs, custom NVHBM memory, and NVLink Fusion into AWS's stack ties Amazon's own Trainium silicon more tightly to NVIDIA's platform rather than positioning it as a standalone alternative, reinforcing NVIDIA as the default substrate for agentic and physical AI workloads even at a hyperscaler with in-house chips. The government AI factory also places NVIDIA and AWS at the center of federal AI procurement, a segment where security-clearance requirements have historically slowed vendor entry.

For builders, RTX PRO 4500-equipped G7 instances (NVIDIA cites 4.6x inference and 2.1x graphics performance over G6) and broader Nemotron availability on Bedrock lower the barrier to running agentic and robotics workloads inside AWS's managed environment — Amazon Robotics is adopting NVIDIA's physical AI platform for warehouse automation. For investors, the size of this commitment, even without a disclosed dollar figure, underscores that GPU capacity remains the binding constraint hyperscalers are racing to relax through 2028.

#NVIDIA #AWS #AIInfrastructure #GPU #ComputeEconomics #PhysicalAI

#NVIDIA#AWS#GPU deployment#Vera CPU#Nemotron#AI infrastructure#related:AWS
Read Original

How This Connects

Based on AI Infra · Player Map

  1. 4h agoNVIDIA and AWS deepen their alliance with 2 million more GPUs, Vera CPU infrastructure, and a government AI factory. · THIS ARTICLE
  2. 1mo agoNVIDIA announced at SIGGRAPH 2026 that its next-generation AI upscaling technology, DLSS 5, will lau...
  3. 1mo agoNVIDIA RTX Spark superchip debuts in laptops at Bilibili World, runs 120B-parameter models locally
  4. 1mo agoNvidia delays next-gen AI rack system Kyber NVL144 by over 12 months to 2028 due to PCB manufacturing issues.

Related News

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard