Google Splits TPU Gen8 into Separate Training and Inference Chips

Google is bifurcating its next-generation TPU into purpose-built training and inference variants — a hardware strategy aimed squarely at NVIDIA's dominance across both workloads.

Google is splitting its eighth-generation TPU into two distinct chip designs: one optimized for training and one for inference, as CNBC reported. The move acknowledges what the industry has learned over the past two years — that the compute profiles for training a frontier model and serving it to millions of users are fundamentally different, and a single chip design forces compromises on both.

Unlock the full briefing

Get every story in today's briefing, the full archive, and the daily AI intelligence brief.

All stories today

Full archive

Daily brief

Cancel anytime. Payments powered by Stripe.