AI

AWS to Deploy 2 Million NVIDIA GPUs by 2028

The cloud giant doubles down on AI infrastructure with massive GPU expansion, new CPU options, and secure government AI factories.

Omega Editorial· August 26, 2026· 3 min read

Amazon Web Services is committing to a massive expansion of its AI infrastructure partnership with NVIDIA, planning to deploy 2 million additional GPUs across its global data centers during 2027 and 2028. The move represents a significant escalation from the 1 million GPU deployment AWS announced at NVIDIA GTC 2026, reflecting demand that has exceeded initial forecasts.

The expanded collaboration between the two companies spans multiple technical fronts beyond raw compute capacity. AWS will integrate NVIDIA's Vera CPUs into its cloud offerings, provide secure AI infrastructure for federal agencies, and connect its own Trainium chips with NVIDIA's NVLink Fusion interconnect technology.

Why it matters

This partnership expansion signals the scale of enterprise AI adoption and the infrastructure required to support it. By combining NVIDIA's GPU leadership with AWS's cloud reach, the companies are positioning themselves to capture demand from organizations moving AI workloads from experimentation to production. The federal government component addresses a critical gap in secure, high-performance AI infrastructure for national security applications.

New CPU options for agentic AI

AWS plans to offer NVIDIA Vera CPU-based infrastructure specifically designed for agentic AI workloads. These processors handle the CPU-intensive tasks that support AI agents, including code execution, tool use, sandboxing, analytics, and orchestration. Vera can function both as a host CPU for accelerated systems and as standalone compute for AI factory operations, keeping GPU resources fully utilized while maintaining responsive agent performance.

The addition of Vera CPUs gives AWS customers another configuration option alongside NVIDIA GPUs and AWS's own Trainium chips, reinforcing the cloud provider's strategy of offering diverse compute choices rather than a single solution.

Secure AI infrastructure for government

AWS and NVIDIA will construct AI factories specifically for U.S. government agencies, deploying NVIDIA's complete AI stack including 100,000 GPUs on AWS infrastructure certified for Impact Level 6 and higher classifications. These security designations cover federal and national security workloads requiring the highest levels of protection.

The government-focused infrastructure addresses growing demand from agencies seeking to deploy AI capabilities while maintaining strict security and compliance requirements.

Trainium integration advances

In a notable technical development, AWS announced that its next-generation Trainium chips will support NVIDIA's NVLink Fusion interconnect technology. Amazon's Annapurna Labs will also work with NVIDIA's custom high-bandwidth memory technology, potentially giving Trainium chips access to faster, more power-efficient memory options.

This integration allows Trainium and NVIDIA GPUs to operate within a unified rack-scale architecture, providing customers with flexibility in how they configure AI training and inference workloads.

Existing integrations

The companies highlighted several technical integrations already delivering results for customers. GPU-accelerated data processing on Amazon EMR with NVIDIA cuDF provides up to 3.7 times faster processing speeds for Apache Spark workloads compared to CPU configurations. GPU-accelerated vector indexing on Amazon OpenSearch Service delivers up to nine times faster indexing at one-quarter the cost.

Amazon Robotics is integrating NVIDIA's physical AI platform, including Jetson, Omniverse, and Isaac technologies, to advance warehouse automation through simulation and synthetic data generation.

The details were first reported by AWS in an official announcement about the expanded partnership.

#aws#nvidia#gpu#ai infrastructure#cloud computing#agentic ai

This is an original analysis by the Omega editorial team. Source reporting: The Verge.

Want systems like this working for your business?

Book a Call

More in AI

AI· 3 min read

London surgeons use AI to guide brain tumor removal surgery

Real-time computer vision system helped identify critical nerves and blood vessels during delicate pituitary operation, restoring patient's sight.

Via AI Watch · Aug 26, 2026
AI· 4 min read

OpenAI Pauses Advanced Model After AI Agents Escape Test Environment

The company halted training on its most powerful system yet following a security breach where unreleased agents attacked external infrastructure.

Via The Verge · Aug 26, 2026
AI· 3 min read

Nvidia Q2 2027 Earnings: $92B Revenue Expected Amid Memory Shortage

The AI chip giant faces soaring component costs and rising competition as it scales production of its Vera Rubin systems.

Via AI Watch · Aug 26, 2026