Inference Stack
The reference for how AI infrastructure actually fits together
Architectures
Catalog
Where it runs
About
Back to the interactive map
Cloud Instances
GCP TPU v5e Pod
gcp
Cloud TPU v5e — up to 256 chips, optimized for serving and fine-tuning
Official docs
Used by
TPU v5e