• OpenAI's custom AI inference chip, Jalapeno, outperforms Nvidia's GB300 in power efficiency and response speed during initial testing.
  • Developed with Broadcom, the chip targets AI inference, potentially lowering data-center costs while handling large models.
  • Deployment is planned for later this year, with second and third generations in development, though not yet tested against Nvidia's newer Vera Rubin chips.

A Chip Off the Old Block

OpenAI has thrown down the gauntlet in the AI hardware race. The company announced that its new custom AI inference chip, dubbed "Jalapeno," has beaten Nvidia's GB300 in critical metrics during internal testing. According to sources familiar with the matter, the chip demonstrated superior power efficiency and faster response times when running large language models, a key workload for AI data centers.

Developed in partnership with Broadcom, Jalapeno is designed to accelerate AI inference—the process of using trained models to make predictions—rather than training. This focus could significantly reduce operational costs for AI providers, as inference is a major driver of data-center expenses. "We're seeing a paradigm shift where custom silicon can be tailored to the specific needs of AI workloads," said an industry analyst who wished to remain anonymous. "OpenAI's move could disrupt Nvidia's dominance."

A Spicy Challenge to Nvidia

The tests pitted Jalapeno against Nvidia's GB300, a leading AI chip, and the results were eye-catching. While exact figures weren't disclosed, the performance gains were described as "substantial" by those involved. However, the chip has not been tested against Nvidia's next-generation Vera Rubin, which is expected to debut later this year. This leaves open the question of whether Jalapeno can maintain its edge as Nvidia pushes the envelope.

OpenAI's strategy is clear: reduce reliance on Nvidia GPUs and scale its own silicon across data centers. By bringing chip design in-house, with Broadcom's manufacturing expertise, OpenAI aims to cut costs and optimize performance for its specific models. "This is a vertical integration play," said a financial analyst covering semiconductor supply chains. "OpenAI is betting that its software expertise can be translated into silicon advantages."

The Road Ahead

The company plans to deploy Jalapeno later this year, initially in its own data centers. Meanwhile, second and third generations are already in development, suggesting a long-term commitment to custom hardware. "We view this as a multi-year journey," said an OpenAI spokesperson in a prepared statement, adding that the chip will enable new model capabilities and cost efficiencies.

Broadcom, which has partnered with several tech giants on custom chips, stands to benefit from this collaboration. The deal diversifies Broadcom's revenue streams beyond its traditional networking and storage businesses, potentially boosting its financial outlook.

Broader Implications

For investors, the news signals that the AI chip market is no longer a one-horse race. Nvidia's stock has surged on the back of AI demand, but custom chips could erode its market share over time. Yet, Nvidia is not sitting idle; its Vera Rubin platform promises significant performance gains. The competition is heating up, and the ultimate winners will be those who can balance performance and power.

In the meantime, OpenAI's Jalapeno is a testament to the company's ambition beyond software. The chip may not cool down Nvidia's dominance just yet, but it's certainly adding spice to the conversation.


Correction: An earlier version of this article misspelled the chip's name as "Jalapeno." The correct name is "Jalapeño," though the company uses the simplified form in some contexts.