Inference Stack
The reference for how AI infrastructure actually fits together
Architectures
Catalog
Where it runs
About
Back to the interactive map
Silicon
TPU v5e
google
Google's 5th-gen TPU — cost-efficient inference and fine-tuning at scale
Official docs
View in catalog
Used by
XLA / JAX
Leads to
GCP TPU v5e Pod