415.tech
AI & tech, from the frontlines of Silicon Valley
OpenAI unveils Jalapeño, its first custom inference chip, co-developed with Broadcom

OpenAI unveils Jalapeño, its first custom inference chip, co-developed with Broadcom

OpenAI unveiled Jalapeño, a custom inference processor co-developed with Broadcom for real-time coding workloads, with early results showing better performance-per-watt than current alternatives — though pre-training workloads will still run on Nvidia. With this move, OpenAI joins Google (TPU), Amazon (Trainium/Inferentia), and Microsoft (Maia) as a frontier-lab operator with proprietary inference silicon, directly pressuring Nvidia's pricing power in the AI accelerator market.

Source: techcrunch.com

Post on XEmail

We have a deep understanding of the workload. We've really been looking for specific workloads that are underserved, [and asking] how can we build something that will be able to accelerate what's possible?

Greg Brockman, OpenAI president

Why this matters

  • → Reduces inference costs, improving AI service profitability at scale
  • → Breaks Nvidia's stranglehold on custom silicon for frontier labs
  • → Enables full-stack optimization across models, hardware, and infrastructure
Breaking Nvidia's grip