Inference Stack
The reference for how AI infrastructure actually fits together
Architectures
Catalog
Where it runs
About
Back to the interactive map
Hardware Abstraction
AWS Neuron SDK
aws
AWS's compiler and runtime SDK for Trainium and Inferentia, integrating with PyTorch and JAX
Official docs
Used by
Triton Inference Server
Leads to
Trainium2
Inferentia2