...
NVIDIA H200 shipments delayed to Q3  · BREAKING: Microsoft confirms 3GW data centre expansion in Asia-Pacific ·  AWS announces new sovereign cloud regions in India and UAE  · Arm-based servers now 24% of hyperscale deployments ·  EU AI Act enforcement enters phase two  · Global data centre investment hits $612B in 2026 ·  TSMC Arizona yields improve to 68% on 3nm process  · OpenAI valuation reaches $400B after latest funding round ·  NVIDIA H200 shipments delayed to Q3  · BREAKING: Microsoft confirms 3GW data centre expansion in Asia-Pacific ·  AWS announces new sovereign cloud regions in India and UAE  · Arm-based servers now 24% of hyperscale deployments ·  EU AI Act enforcement enters phase two  · Global data centre investment hits $612B in 2026
NVIDIA H200 shipments delayed to Q3  · BREAKING: Microsoft confirms 3GW data centre expansion in Asia-Pacific ·  AWS announces new sovereign cloud regions in India and UAE  · Arm-based servers now 24% of hyperscale deployments ·  EU AI Act enforcement enters phase two  · Global data centre investment hits $612B in 2026 ·  TSMC Arizona yields improve to 68% on 3nm process  · OpenAI valuation reaches $400B after latest funding round ·  NVIDIA H200 shipments delayed to Q3  · BREAKING: Microsoft confirms 3GW data centre expansion in Asia-Pacific ·  AWS announces new sovereign cloud regions in India and UAE  · Arm-based servers now 24% of hyperscale deployments ·  EU AI Act enforcement enters phase two  · Global data centre investment hits $612B in 2026

AWS expands Nvidia GPU deployment for AI infrastructure

Amazon Web Services is preparing for another major expansion of its AI computing infrastructure, committing to deploy an additional 2

Share
AWS GPU expansion

Amazon Web Services is preparing for another major expansion of its AI computing infrastructure, committing to deploy an additional 2 million Nvidia GPUs across its global cloud footprint. The deployment will span 2027 and 2028, extending an already aggressive buildout that reflects how quickly demand for accelerated computing has moved beyond earlier capacity forecasts. The agreement also broadens the relationship between AWS and Nvidia beyond GPUs, covering CPUs, networking, open models, data processing, and robotics. For AWS, the move positions Nvidia infrastructure as a deeper part of its long-term strategy for serving frontier AI developers, enterprises, and government customers.

The new GPU commitment includes Nvidia’s Blackwell Ultra, Rubin, and Rubin Ultra platforms, giving AWS a pipeline that stretches across multiple generations of accelerator technology. AWS had previously said during Nvidia GTC 2026 that it planned to bring more than one million Nvidia GPUs online across its platform during 2026. Demand has since exceeded that expectation, pushing the companies toward a substantially larger deployment schedule. The scale of the latest commitment also suggests that hyperscale cloud providers are planning capacity further ahead as AI training, inference, and agentic workloads continue to consume increasingly large amounts of compute.

AWS expands beyond GPUs into the full AI stack

The agreement reaches into the processor layer as AWS prepares to introduce Nvidia Vera CPU-based infrastructure on its cloud platform. The companies also plan to extend Nvidia NVLink Fusion with custom Nvidia high-bandwidth memory, creating another path for tightly integrated compute architectures as AI systems become more demanding. AWS will build AI factories for the US government using 100,000 GPUs on secure AWS infrastructure, adding a significant government-focused component to the collaboration. The broader arrangement points toward infrastructure designed around complete AI systems rather than isolated accelerator deployments.

AWS is also adding Nvidia’s Nemotron models to its platform, widening the software choices available to customers building AI applications. The partnership will cover data processing and open-model capabilities alongside the underlying compute, networking, and security layers. Amazon Robotics will further work with Nvidia to adopt the company’s physical AI platform, connecting cloud-scale AI infrastructure with robotics systems operating in physical environments. At the same time, the partnership is becoming a platform strategy that links model development, infrastructure, deployment, and physical AI rather than treating each area as a separate investment.

Blackwell capacity expands across EC2

AWS is also increasing Blackwell capacity through Nvidia RTX Pro 4500 Blackwell Server Edition GPUs for its EC2 G7 instances. AWS says the instances deliver 4.6 times the AI inference performance and 2.1 times the graphics performance of previous-generation G6 instances. The move gives customers another option for workloads that require a combination of accelerated inference and graphics capabilities without relying solely on the highest-end training systems. It also shows how AWS is expanding the Nvidia portfolio across different performance tiers as cloud demand becomes more varied.

The infrastructure push comes as Nvidia continues to report exceptional demand for its computing platforms. Nvidia reported $96.2 billion in revenue for the second quarter of fiscal 2026, representing an 18 percent increase from the previous quarter and a 106 percent increase from the same period a year earlier. Both GAAP and non-GAAP gross margins reached 75 percent during the quarter, underscoring the financial strength behind the accelerator market. However, AWS’s decision to add two million more GPUs indicates that customer demand remains a central constraint even as Nvidia scales production and introduces new generations.

AWS keeps older Nvidia systems in service

AWS’s strategy also includes extending the useful life of existing Nvidia hardware rather than treating each new generation as an immediate replacement cycle. CEO Matt Garman has previously said AWS is still operating six-year-old Nvidia A100 servers and has “never retired an A100” server. That approach gives AWS another lever for managing capacity as newer Blackwell and Rubin systems enter the fleet. It also highlights the importance of total compute availability, where older accelerators can continue serving suitable workloads while newer systems handle more demanding AI applications.

“Nvidia and AWS have built one of the great growth engines of the AI era, and demand is running ahead of every forecast,” said Jensen Huang, founder and CEO of Nvidia. “For 16 years, we have scaled Nvidia computing in the cloud together. Now, we are expanding our partnership across the full stack — GPUs, CPUs, networking, open models and software — to make agentic and physical AI real at an unprecedented pace and scale that only AWS and Nvidia can deliver. This expansion reflects customers’ demand for Nvidia’s platform on AWS.”

Cloud AI capacity becomes a strategic battleground

The latest commitment changes the scale at which AWS is planning its Nvidia relationship. Two million additional GPUs represent not just a hardware purchase but a multi-year bet on sustained demand for accelerated computing across commercial AI, inference, government systems, robotics, and emerging agentic workloads. The inclusion of CPUs, networking, memory, models, and physical AI suggests that AWS wants customers to consume an increasingly integrated Nvidia technology stack through its cloud. That strategy could strengthen AWS’s position among organizations that want Nvidia’s newest platforms without building and operating the corresponding infrastructure themselves.

“Customers want the freedom to choose the best tools for their AI workloads, and they want confidence that everything works seamlessly together,” said Matt Garman, CEO of AWS. “That’s why we’ve invested deeply with Nvidia to make AWS the best place to run Nvidia AI technologies, optimizing performance across our infrastructure from networking and security to deployment. This expanded collaboration gives frontier labs, enterprises and governments even more ways to build and deploy AI on AWS.”

The AWS-Nvidia agreement ultimately reflects a broader shift in the economics of cloud infrastructure, where access to GPUs increasingly determines how quickly customers can train, deploy, and scale AI systems. AWS is responding by committing capacity years ahead while supporting several generations of Nvidia accelerators and keeping older hardware productive inside its fleet. The strategy also gives Nvidia a deeper route into cloud deployments across compute, software, networking, and physical AI. With two million more GPUs scheduled for 2027 and 2028, the companies are effectively betting that the current AI infrastructure race has much further to run.

[simple-author-box]

More from AI Infrastructure

MTN Group is deepening its data center ambitions through a new partnership with a

Eurofiber has secured €2.2 billion in long-term financing through a sustainability-linked refinancing, giving the

India’s accelerating data centre construction cycle is creating a widening market for companies that

COMPUTE WEEKLY

The briefing that 40,000+ tech leaders read every Monday. Sharp, fast, essential.

Great! We’ve received your information.

Building an AI Startup Without Owning GPUs

Not owning GPUs has become the default, deliberate strategy for building an AI company — not a compromise founders accept reluctantly. H100 rental rates fell 64-75% in fifteen months, a dense ecosystem of neoclouds and inference-as-a-service providers now lets startups skip infrastructure entirely, and credit programs can fund a company’s first year before a founder writes a check
Most Read

AI infrastructure decisions for high-density deployments increasingly involve what happens after electricity enters the

Demand is broadening across enterprise workloads APAC’s infrastructure story is changing in ways that

AI infrastructure decisions increasingly influence what enterprises can build, test, and deliver. They also

Why Infrastructure Planning Now Starts With Availability A data center project can have a

A property can look enormous from the site entrance and still offer almost no

Disruptor Spotlight

Cerebras Systems

The chip that makes Nvidia nervous. Cerebras’ Wafer Scale Engine is rewriting the rules of AI inference at scale.
Faster
0 x
YoY Revenue
0 x
Transistors
0 T
Market Pulse
MSFT
+1.02%
NVDA
+0.66%
AMZN
-0.078%
AMD
-6.95%
TSMC
-2.98%
Indicative only · Not financial advice
Upcoming Events
SEP
The AI Infrastructure Race (India)
WEBINAR · ONLINE
The AI Infrastructure Race: Won on Power, Land and Trust — Not Capital
MAY
0
AI Infrastructure Summit
DUBAI · IN PERSON
MEA’s premier AI infrastructure event.
JUN
0 0
Compute Forecast Summit
SINGAPORE · IN PERSON
Our flagship APAC event. Early bird open.
Latest Moves
Live
ecolab
Ecolab Deepens Cooling Strategy With $4.75B CoolIT Acquisition
Ecolab is making one of its biggest moves yet into AI infrastructure after completing its $4.75 billion acquisition of liquid cooling specialist CoolIT Systems
Pure DC AVK Europe data center microgrid Dublin 110MW AI infrastructure Ireland 2026
Pure DC and AVK Deploy Europe’s First 110 MW Data Center Microgrid in Dublin
The Pure DC Dublin microgrid has made history as Europe’s first large-scale on-site data center microgrid, launched in partnership with power solutions provider AVK at Pure DC’s campus in Ireland.
Pace Digitek
Pace Digitek Partners With MEGMEET to Expand AI Data Center Power Business
India’s AI infrastructure ecosystem continues to mature as domestic technology manufacturers move beyond traditional telecommunications and industrial markets toward high-growth digital infrastructure opportunities
Follow Compute Forecast
11K followers
1200 followers
Companies to Watch
CW
CoreWeave
Neo Cloud · $19B · IPO Watch
CB
Cerebras Systems
AI Hardware · $4.25B · Pre-IPO
G42
G42
Sovereign AI · Abu Dhabi
H
Humain
Saudi AI · $40B Fund
Latest Podcast
AI Capex, Cloud Margins & the Nuclear Bet
48 MIN · 25 APR 2026

AWS expands Nvidia GPU deployment for AI infrastructure

Amazon Web Services is preparing for another major expansion of its AI computing infrastructure, committing to deploy an additional 2

Share
AWS GPU expansion
2
847 SHARES

0
SHARES

[simple-author-box]

More from AI Infrastructure

AI infrastructure decisions for high-density deployments increasingly involve what happens after electricity enters the

Demand is broadening across enterprise workloads APAC’s infrastructure story is changing in ways that

AI infrastructure decisions increasingly influence what enterprises can build, test, and deliver. They also

Why Infrastructure Planning Now Starts With Availability A data center project can have a

COMPUTE WEEKLY

The briefing that 40,000+ tech leaders read every Monday. Sharp, fast, essential.

Great! We’ve received your information.

Global AI Infrastructure Outlook 2026

The briefing that 40,000+ tech leaders read every Monday. Sharp, fast, essential.
Download Free
Most Read

AI infrastructure decisions for high-density deployments increasingly involve what happens after electricity enters the

Demand is broadening across enterprise workloads APAC’s infrastructure story is changing in ways that

AI infrastructure decisions increasingly influence what enterprises can build, test, and deliver. They also

Why Infrastructure Planning Now Starts With Availability A data center project can have a

A property can look enormous from the site entrance and still offer almost no

Disruptor Spotlight

Cerebras Systems

The chip that makes Nvidia nervous. Cerebras’ Wafer Scale Engine is rewriting the rules of AI inference at scale.
Faster
0 x
YoY Revenue
0 x
Transistors
0 T
Market Pulse
NVDA
$924.60
+2.4%
MSFT
$421.30
+1.1%
AMZN
$192.80
-0.6%
NVDA
$924.60
+2.4%
NVDA
$924.60
+2.4%
Indicative only · Not financial advice
Upcoming Events
MAY
0 0
DCD Global — London
LONDON · IN PERSON
World’s largest DC event. CF is media partner.
MAY
0
AI Infrastructure Summit
DUBAI · IN PERSON
MEA’s premier AI infrastructure event.
JUN
0 0

Compute Forecast Summit

SINGAPORE · IN PERSON
Our flagship APAC event. Early bird open.
Latest Moves
  • Live
Sam Altman
OpenAI appoints new Chief Infrastructure Officer to lead $100B DC programme
27 APR · OPENAI
Sam Altman
OpenAI appoints new Chief Infrastructure Officer to lead $100B DC programme
27 APR · OPENAI
Sam Altman
OpenAI appoints new Chief Infrastructure Officer to lead $100B DC programme
27 APR · OPENAI
Follow Compute Forecast
18.4K followers
12.1K followers
9.3K subscribers
41 episodes
Companies to Watch
CW
CoreWeave
Neo Cloud · $19B · IPO Watch
CB
Cerebras Systems
AI Hardware · $4.25B · Pre-IPO
G42
G42
Sovereign AI · Abu Dhabi
CW
Humain
Saudi AI · $40B Fund
Latest Podcast
AI Capex, Cloud Margins & the Nuclear Bet
48 MIN · 25 APR 2026
Scroll to Top
Seraphinite AcceleratorOptimized by Seraphinite Accelerator
Turns on site high speed to be attractive for people and search engines.