Inference Stack
The reference for how AI infrastructure actually fits together
Architectures
Catalog
Where it runs
About
Back to the interactive map
Silicon
H100
nvidia
Hopper-architecture data center GPU — 80 GB HBM3, 3.35 TB/s, 4th-gen NVLink
Official docs
View in catalog
Used by
CUDA
Leads to
AWS p5 (H100)
GCP a3-mega (H100)