...
NVIDIA H200 shipments delayed to Q3  · BREAKING: Microsoft confirms 3GW data centre expansion in Asia-Pacific ·  AWS announces new sovereign cloud regions in India and UAE  · Arm-based servers now 24% of hyperscale deployments ·  EU AI Act enforcement enters phase two  · Global data centre investment hits $612B in 2026 ·  TSMC Arizona yields improve to 68% on 3nm process  · OpenAI valuation reaches $400B after latest funding round ·  NVIDIA H200 shipments delayed to Q3  · BREAKING: Microsoft confirms 3GW data centre expansion in Asia-Pacific ·  AWS announces new sovereign cloud regions in India and UAE  · Arm-based servers now 24% of hyperscale deployments ·  EU AI Act enforcement enters phase two  · Global data centre investment hits $612B in 2026
NVIDIA H200 shipments delayed to Q3  · BREAKING: Microsoft confirms 3GW data centre expansion in Asia-Pacific ·  AWS announces new sovereign cloud regions in India and UAE  · Arm-based servers now 24% of hyperscale deployments ·  EU AI Act enforcement enters phase two  · Global data centre investment hits $612B in 2026 ·  TSMC Arizona yields improve to 68% on 3nm process  · OpenAI valuation reaches $400B after latest funding round ·  NVIDIA H200 shipments delayed to Q3  · BREAKING: Microsoft confirms 3GW data centre expansion in Asia-Pacific ·  AWS announces new sovereign cloud regions in India and UAE  · Arm-based servers now 24% of hyperscale deployments ·  EU AI Act enforcement enters phase two  · Global data centre investment hits $612B in 2026

Google I/O 2026: TPU 8i and 8t Signal a New Era for AI Infrastructure

Google confirmed at I/O 2026 on May 19 that it expects to spend approximately $180 to $190 billion in capital

Share
Google IO TPU 8i 8t new AI infrastructure era 2026 dual chip training inference Sundar Pichai capex

Google confirmed at I/O 2026 on May 19 that it expects to spend approximately $180 to $190 billion in capital expenditure this year, roughly six times the $31 billion it spent in 2022. The centrepiece of that investment is a dual-chip eighth-generation TPU architecture that Google described as the most significant infrastructure milestone since its first commercial TPU in 2016. For the first time, Google has split its TPU architecture into two purpose-built chips: TPU 8t for large-scale model training and TPU 8i for inference workloads. The training chip scales across more than one million TPUs globally through JAX and Pathways, forming what Google calls the largest training cluster in the world.

The inference chip triples on-chip SRAM to 384MB, increases high-bandwidth memory to 288GB, doubles ICI bandwidth to 19.2 terabits per second, and delivers 1,500 tokens per second on production inference workloads — 80% better performance per dollar than the prior generation.

What the Dual-Chip Architecture Signals

The decision to split training and inference into separate silicon is the most consequential infrastructure architecture decision Google has made since it first built TPUs to replace GPUs for internal workloads. Training and inference have fundamentally different requirements. Training requires raw compute throughput and high inter-chip bandwidth across massive synchronised clusters. Inference requires low latency, high memory bandwidth for KV cache serving, and cost-efficient token throughput at scale. A single chip optimised for both produces compromises on each. Two purpose-built chips produce the best possible performance on each at the cost of managing two hardware ecosystems. Google’s willingness to absorb that operational complexity signals its confidence that the inference workload is large enough, permanent enough, and economically distinct enough from training to justify dedicated silicon.

The Anthropic-AWS $100 billion compute deal documented that the frontier AI lab compute market is consolidating around decade-long infrastructure commitments. The TPU 8i is Google’s answer to the question of what infrastructure those commitments will run on.

What This Means for Nvidia and the Infrastructure Market

The TPU 8i directly addresses the inference market where Google-Blackstone’s new TPU cloud venture competes with Nvidia-backed neoclouds. 80% better performance per dollar for inference is a commercially significant claim at a moment when enterprise AI budgets are 85% inference spend. If the TPU 8i delivers on that claim at production scale, it validates the Google-Blackstone joint venture’s commercial case and provides the competitive infrastructure cost advantage that makes the venture viable against Nvidia’s GPU ecosystem. The infrastructure market may now have a second credible roadmap to evaluate alongside Nvidia’s, with $180 to $190 billion in annual Google capex ensuring that roadmap will be executed regardless of supply chain constraints or competitive pressure.

[simple-author-box]

More from AI Infrastructure

Modus has raised $10 million in seed funding to build infrastructure designed to solve

CVC DIF is moving deeper into Germany’s digital infrastructure market with an agreement to

Larsen & Toubro (L&T), through its AI infrastructure businesses, has secured a major contract

COMPUTE WEEKLY

The briefing that 40,000+ tech leaders read every Monday. Sharp, fast, essential.

Great! We’ve received your information.

Building an AI Startup Without Owning GPUs

Not owning GPUs has become the default, deliberate strategy for building an AI company — not a compromise founders accept reluctantly. H100 rental rates fell 64-75% in fifteen months, a dense ecosystem of neoclouds and inference-as-a-service providers now lets startups skip infrastructure entirely, and credit programs can fund a company’s first year before a founder writes a check
Most Read

AI infrastructure decisions increasingly influence what enterprises can build, test, and deliver. They also

Why Infrastructure Planning Now Starts With Availability A data center project can have a

A property can look enormous from the site entrance and still offer almost no

As rack power rises toward the megawatt range, the physical footprint of power-delivery equipment

A data center project can look complete long before it delivers usable capacity. The

Disruptor Spotlight

Cerebras Systems

The chip that makes Nvidia nervous. Cerebras’ Wafer Scale Engine is rewriting the rules of AI inference at scale.
Faster
0 x
YoY Revenue
0 x
Transistors
0 T
Market Pulse
MSFT
+1.02%
NVDA
+0.66%
AMZN
-0.078%
AMD
-6.95%
TSMC
-2.98%
Indicative only · Not financial advice
Upcoming Events
SEP
The AI Infrastructure Race (India)
WEBINAR · ONLINE
The AI Infrastructure Race: Won on Power, Land and Trust — Not Capital
MAY
0
AI Infrastructure Summit
DUBAI · IN PERSON
MEA’s premier AI infrastructure event.
JUN
0 0
Compute Forecast Summit
SINGAPORE · IN PERSON
Our flagship APAC event. Early bird open.
Latest Moves
Live
ecolab
Ecolab Deepens Cooling Strategy With $4.75B CoolIT Acquisition
Ecolab is making one of its biggest moves yet into AI infrastructure after completing its $4.75 billion acquisition of liquid cooling specialist CoolIT Systems
Pure DC AVK Europe data center microgrid Dublin 110MW AI infrastructure Ireland 2026
Pure DC and AVK Deploy Europe’s First 110 MW Data Center Microgrid in Dublin
The Pure DC Dublin microgrid has made history as Europe’s first large-scale on-site data center microgrid, launched in partnership with power solutions provider AVK at Pure DC’s campus in Ireland.
Pace Digitek
Pace Digitek Partners With MEGMEET to Expand AI Data Center Power Business
India’s AI infrastructure ecosystem continues to mature as domestic technology manufacturers move beyond traditional telecommunications and industrial markets toward high-growth digital infrastructure opportunities
Follow Compute Forecast
11K followers
1200 followers
Companies to Watch
CW
CoreWeave
Neo Cloud · $19B · IPO Watch
CB
Cerebras Systems
AI Hardware · $4.25B · Pre-IPO
G42
G42
Sovereign AI · Abu Dhabi
H
Humain
Saudi AI · $40B Fund
Latest Podcast
AI Capex, Cloud Margins & the Nuclear Bet
48 MIN · 25 APR 2026

Google I/O 2026: TPU 8i and 8t Signal a New Era for AI Infrastructure

Google confirmed at I/O 2026 on May 19 that it expects to spend approximately $180 to $190 billion in capital

Share
Google IO TPU 8i 8t new AI infrastructure era 2026 dual chip training inference Sundar Pichai capex
158
847 SHARES

0
SHARES

[simple-author-box]

More from AI Infrastructure

AI infrastructure decisions increasingly influence what enterprises can build, test, and deliver. They also

Why Infrastructure Planning Now Starts With Availability A data center project can have a

A property can look enormous from the site entrance and still offer almost no

As rack power rises toward the megawatt range, the physical footprint of power-delivery equipment

COMPUTE WEEKLY

The briefing that 40,000+ tech leaders read every Monday. Sharp, fast, essential.

Great! We’ve received your information.

Global AI Infrastructure Outlook 2026

The briefing that 40,000+ tech leaders read every Monday. Sharp, fast, essential.
Download Free
Most Read

AI infrastructure decisions increasingly influence what enterprises can build, test, and deliver. They also

Why Infrastructure Planning Now Starts With Availability A data center project can have a

A property can look enormous from the site entrance and still offer almost no

As rack power rises toward the megawatt range, the physical footprint of power-delivery equipment

A data center project can look complete long before it delivers usable capacity. The

Disruptor Spotlight

Cerebras Systems

The chip that makes Nvidia nervous. Cerebras’ Wafer Scale Engine is rewriting the rules of AI inference at scale.
Faster
0 x
YoY Revenue
0 x
Transistors
0 T
Market Pulse
NVDA
$924.60
+2.4%
MSFT
$421.30
+1.1%
AMZN
$192.80
-0.6%
NVDA
$924.60
+2.4%
NVDA
$924.60
+2.4%
Indicative only · Not financial advice
Upcoming Events
MAY
0 0
DCD Global — London
LONDON · IN PERSON
World’s largest DC event. CF is media partner.
MAY
0
AI Infrastructure Summit
DUBAI · IN PERSON
MEA’s premier AI infrastructure event.
JUN
0 0

Compute Forecast Summit

SINGAPORE · IN PERSON
Our flagship APAC event. Early bird open.
Latest Moves
  • Live
Sam Altman
OpenAI appoints new Chief Infrastructure Officer to lead $100B DC programme
27 APR · OPENAI
Sam Altman
OpenAI appoints new Chief Infrastructure Officer to lead $100B DC programme
27 APR · OPENAI
Sam Altman
OpenAI appoints new Chief Infrastructure Officer to lead $100B DC programme
27 APR · OPENAI
Follow Compute Forecast
18.4K followers
12.1K followers
9.3K subscribers
41 episodes
Companies to Watch
CW
CoreWeave
Neo Cloud · $19B · IPO Watch
CB
Cerebras Systems
AI Hardware · $4.25B · Pre-IPO
G42
G42
Sovereign AI · Abu Dhabi
CW
Humain
Saudi AI · $40B Fund
Latest Podcast
AI Capex, Cloud Margins & the Nuclear Bet
48 MIN · 25 APR 2026
Scroll to Top
Seraphinite AcceleratorOptimized by Seraphinite Accelerator
Turns on site high speed to be attractive for people and search engines.