Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. Google unveils eighth-generation TPUs, split into TPU 8t…

Google unveils eighth-generation TPUs, split into TPU 8t (training) and TPU 8i (inference)

★★★hardware-computeGoogleconfidence: medium

At Google Cloud Next 2026 (April) Google announced its first split TPU generation: TPU 8t for training (pods of 9,600 chips, 2 PB shared memory, 121 exaFLOPS) and TPU 8i for inference (288 GB HBM, 80% better perf/$), both up to 2x better performance-per-watt than Ironwood, which became generally available at the same event.

Key facts

What happened

Google introduced two purpose-built eighth-generation TPUs at Cloud Next 2026 in Las Vegas, with general availability promised later in 2026 as part of AI Hypercomputer.

Why it matters

Separate training and inference silicon reflects how agentic, long-running inference now dominates compute demand, and strengthens Google's position as the main non-NVIDIA accelerator supplier (Anthropic is reported as an anchor customer).

Changelog

  • 2026-09-29: created (exact announcement day inferred from press dated 2026-04-22; confidence medium)

Related events

  1. Alphabet Q2 2026: Google Cloud +82%, capex guidance raised to up to $205B, Gemini at 22B API tokens/minute ★★★

Sources (3)

id: 2026-04-22-google-tpu-8t-8i · updated 2026-09-29 · open in the interactive timeline