AI

AWS to Deploy 2 Million NVIDIA GPUs by 2028

The cloud giant doubles down on AI infrastructure with massive GPU expansion, new CPU options, and secure government AI factories.

Omega Editorial· August 26, 2026· 3 min read

Amazon Web Services is committing to a massive expansion of its AI infrastructure partnership with NVIDIA, planning to deploy 2 million additional GPUs across its global data centers during 2027 and 2028. The move represents a significant escalation from the 1 million GPU deployment AWS announced at NVIDIA GTC 2026, reflecting demand that has exceeded initial forecasts.

The expanded collaboration between the two companies spans multiple technical fronts beyond raw compute capacity. AWS will integrate NVIDIA's Vera CPUs into its cloud offerings, provide secure AI infrastructure for federal agencies, and connect its own Trainium chips with NVIDIA's NVLink Fusion interconnect technology.

Why it matters

This partnership expansion signals the scale of enterprise AI adoption and the infrastructure required to support it. By combining NVIDIA's GPU leadership with AWS's cloud reach, the companies are positioning themselves to capture demand from organizations moving AI workloads from experimentation to production. The federal government component addresses a critical gap in secure, high-performance AI infrastructure for national security applications.

New CPU options for agentic AI

AWS plans to offer NVIDIA Vera CPU-based infrastructure specifically designed for agentic AI workloads. These processors handle the CPU-intensive tasks that support AI agents, including code execution, tool use, sandboxing, analytics, and orchestration. Vera can function both as a host CPU for accelerated systems and as standalone compute for AI factory operations, keeping GPU resources fully utilized while maintaining responsive agent performance.

The addition of Vera CPUs gives AWS customers another configuration option alongside NVIDIA GPUs and AWS's own Trainium chips, reinforcing the cloud provider's strategy of offering diverse compute choices rather than a single solution.

Secure AI infrastructure for government

AWS and NVIDIA will construct AI factories specifically for U.S. government agencies, deploying NVIDIA's complete AI stack including 100,000 GPUs on AWS infrastructure certified for Impact Level 6 and higher classifications. These security designations cover federal and national security workloads requiring the highest levels of protection.

The government-focused infrastructure addresses growing demand from agencies seeking to deploy AI capabilities while maintaining strict security and compliance requirements.

Trainium integration advances

In a notable technical development, AWS announced that its next-generation Trainium chips will support NVIDIA's NVLink Fusion interconnect technology. Amazon's Annapurna Labs will also work with NVIDIA's custom high-bandwidth memory technology, potentially giving Trainium chips access to faster, more power-efficient memory options.

This integration allows Trainium and NVIDIA GPUs to operate within a unified rack-scale architecture, providing customers with flexibility in how they configure AI training and inference workloads.

Existing integrations

The companies highlighted several technical integrations already delivering results for customers. GPU-accelerated data processing on Amazon EMR with NVIDIA cuDF provides up to 3.7 times faster processing speeds for Apache Spark workloads compared to CPU configurations. GPU-accelerated vector indexing on Amazon OpenSearch Service delivers up to nine times faster indexing at one-quarter the cost.

Amazon Robotics is integrating NVIDIA's physical AI platform, including Jetson, Omniverse, and Isaac technologies, to advance warehouse automation through simulation and synthetic data generation.

The details were first reported by AWS in an official announcement about the expanded partnership.

#aws#nvidia#gpu#ai infrastructure#cloud computing#agentic ai

This is an original analysis by the Omega editorial team. Source reporting: The Verge.

Want systems like this working for your business?

Book a Call

More in AI

AI· 3 min read

Google adds encrypted cloud memory to Private AI Compute

New architecture lets AI assistants retain context across devices while keeping data inaccessible to Google itself through device-held encryption keys.

Via AI Watch · Sep 24, 2026
AI· 3 min read

Zuckerberg Says AI Outpaced Metaverse Hardware, Prompting Shift

Meta's CEO acknowledges the company pivoted strategy after artificial intelligence capabilities advanced faster than affordable holographic technology.

Via AI Watch · Sep 24, 2026
AI· 4 min read

Computer Science Grads Pivot to AI Roles as Entry-Level Coding Jobs Vanish

Recent graduates face a 7.1% unemployment rate as tech giants automate development work and smaller firms seek AI implementation help instead.

Via AI Watch · Sep 24, 2026