Skip to content
Tech News
← Back to articles

OpenAI’s Jalapeño AI chip brings new 'threat' to Nvidia margins as custom silicon gains ground

read original more articles
Why This Matters

OpenAI's introduction of the Jalapeño AI chip signifies a shift in the industry, challenging Nvidia's dominant position in AI hardware with custom-designed chips that offer comparable or superior inference efficiency. This development could lead to increased competition and innovation in AI chip manufacturing, impacting both industry giants and consumers by potentially lowering costs and improving AI performance. The move highlights a growing trend of hyperscalers developing proprietary silicon to optimize AI workloads.

Key Takeaways

Nvidia 's near-monopoly over the most advanced AI chips is under "threat" as OpenAI's and other tech giants announce custom-built semiconductors, analysts told CNBC.

OpenAI announced that its first AI chip, the Jalapeño, had "industry-leading speed and efficiency," as it unveiled the semiconductor on Tuesday. Google, AWS and Meta are all also developing their own AI chips.

Nvidia has seen its share price rocket amid the data center buildout, which has created huge demand for its chips in both model training and inference: how AI systems run day-to-day tasks.

But hyperscalers and AI companies are increasingly gaining ground in developing their own silicon to power AI systems.

The Jalapeño chip, designed for inference, shows that a "hyperscaler-designed chip can now match or beat Nvidia's Blackwell-class GPUs on inference efficiency," Adrien Sanchez, technology analyst at Yole Group, told CNBC.

He added that, while Nvidia still owns the "vast majority" of AI compute and has ecosystem lock-in to its software platform CUDA, OpenAI's new chip is a "threat to Nvidia's inference margins, which is the field growing the most at the moment."

Nvidia has been approached for comment.

OpenAI announced the first benchmarking results from its Jalapeño, saying it would allow users to get "faster responses, more responsive agents and more reliable access" as demand grows.

The new chip is being developed with Broadcom . It will be deployed within OpenAI's compute infrastructure by the end of the year, and OpenAI said it was already working on the semiconductor's generations two and three.