Inference Stack
The reference for how AI infrastructure actually fits together
Architectures
Catalog
Where it runs
About
Back to the interactive map
Hardware Abstraction
XLA / JAX
google
Accelerated Linear Algebra compiler used by JAX and TensorFlow for TPUs
Official docs
Used by
Ray Serve
Leads to
TPU v5e