AI inference costNVIDIA Alternatives for Inference: Only Trainium Pays Off
Normalised to dollars per TB/s of memory bandwidth per hour, only AWS Trainium2 beats NVIDIA on a price you can actually get quoted. Google's best inference chip has no published rate, AMD's case dies at hyperscaler pricing, and Intel Gaudi 3 has no successor.
September 3, 2026 · 17 min readOptimumNvidia Buys the Bridge to Trainium. Go Grep Your Imports.
Nvidia's reported $12.9B deal for Hugging Face puts optimum-neuron, optimum-habana, optimum-intel and optimum-amd — the Transformers bridge for every rival accelerator — inside Nvidia. optimum-tpu was archived read-only in January, and its own README points the exit at vLLM.
August 27, 2026 · 10 min read