Where it runs
NVIDIA GPUs and each cloud's own silicon (AWS Trainium/Inferentia, Google TPU) traced through to the instance families built on them.
NVIDIA GPUs and each cloud's own silicon (AWS Trainium/Inferentia, Google TPU) traced through to the instance families built on them.
First cloud in the 'where this hardware actually shows up' series. EC2 instance family → the silicon inside it (NVIDIA and AWS's own Trainium/Inferentia), plus the software integration points and what else on AWS is neither.
NVIDIA's managed AI infrastructure service, sold and run on AWS capacity.
NIM microservices (including NVIDIA Nemotron open models and Cosmos world models) deployable through multiple AWS services.
AWS's new sovereign-AI infrastructure offering combines AWS services with NVIDIA Blackwell GPUs and Spectrum-X networking.
AWS's own next-gen custom silicon (Trainium4) is being designed to interconnect via NVIDIA NVLink and the MGX rack architecture — NVIDIA's interconnect reaching beyond NVIDIA's own chips.
NVIDIA's GPU-accelerated vector search library now powers OpenSearch Serverless vector search.
The host/hypervisor offload layer on every modern EC2 GPU instance is AWS's own Nitro System, not NVIDIA BlueField — a common point of confusion.