Inference Stack
The reference for how AI infrastructure actually fits together
Architectures
Catalog
Where it runs
About
Back to the interactive map
Silicon
L4
nvidia
Ada Lovelace inference GPU — 24 GB GDDR6, 72 W TDP, cost-efficient inference
Official docs
View in catalog
Used by
CUDA
Leads to
AWS g6 (L4)
GCP g2 (L4)