
Qualcomm will design custom AI inference chips for AWS, its third data-center win since June
Qualcomm will design custom AI inference silicon for AWS across multiple chip generations, plus optical interconnects up to 1.6 terabits, aimed at cutting energy cost per token — and uses Amazon Bedrock to speed its own chip design in return. It is Qualcomm's third data-center win since June, after Meta's Dragonfly C1000 and Microsoft's HBC memory in Azure, giving weight to its $15B data-center revenue target for 2029. The chips sit alongside AWS's own Trainium, Graviton, and Nitro families, turning inference — where cost per token drives margins — into a multi-vendor race inside AWS itself.
Source: the-decoder.com ↗
The move targets inference workloads, where energy cost per token drives the bottom line.
the-decoder.com
Why this matters
- → Inference cost per token is becoming the margin battle inside major cloud providers
- → Qualcomm's third major win in four months signals sustained demand for power-efficient custom silicon
- → Cloudflare-style vertical integration: AWS designs with Bedrock while Qualcomm designs for AWS
Inference arms race