Inference Stack
The reference for how AI infrastructure actually fits together
Architectures
Catalog
Where it runs
About
Back to the interactive map
Hardware Abstraction
ROCm
amd
AMD's open-source GPU compute platform and HIP programming interface
Official docs
Used by
Triton Inference Server
Ray Serve
SGLang Runtime
Leads to
MI300X