TPU 8t and TPU 8i

Google Cloud has announced its eighth generation of custom Tensor Processor Units (TPUs), introducing two distinct chips purpose-built to handle the unique demands of the "agentic era." As AI moves from single-turn chatbots to autonomous agents that reason, plan, and execute multi-step workflows, Google is separating its infrastructure to optimize for two fundamentally different workloads: Training and Inference.

Two Specialized Chips

Engineering for the Agentic Era

Google has redesigned the TPU stack to eliminate the "waiting room" effect where agents sit idle while waiting for data:

Reliability at Scale

Recognizing that downtime is the enemy of frontier-scale training, TPU 8t is engineered to target over 97% "goodput" (productive compute time). This is achieved through:

Availability and Ecosystem

This launch represents a foundational shift. As businesses move from "simple AI" to "agentic AI" that executes continuous loops of reasoning, the bottleneck shifts from simple raw compute to memory bandwidth and interconnect latency. By specializing silicon for these specific agentic bottlenecks, Google is setting the stage for a new generation of autonomous applications that are as fast as they are capable. Interested developers can request more information on the Google Cloud website.