AI inference costNVIDIA Alternatives for Inference: Only Trainium Pays Off
Normalised to dollars per TB/s of memory bandwidth per hour, only AWS Trainium2 beats NVIDIA on a price you can actually get quoted. Google's best inference chip has no published rate, AMD's case dies at hyperscaler pricing, and Intel Gaudi 3 has no successor.
September 3, 2026 · 17 min readQualcommQualcomm Spent $4B to Break Nvidia's Lock on Enterprise AI
Qualcomm's $3.92 billion acquisition of Modular — maker of the Mojo language and MAX inference engine — is not a chip deal. It's a direct attack on CUDA, the software platform that has locked 4 million developers and their enterprises into Nvidia's ecosystem for nearly two decades. Combined with a reported $8-10 billion Tenstorrent acquisition, Qualcomm is assembling a $14 billion full-stack alternative for the $255 billion AI inference market. Here's how to assess your own Nvidia lock-in and plan a multi-vendor inference strategy.
June 26, 2026 · 17 min read