
OpenAI and Broadcom have unveiled Jalapeño, a custom chip designed specifically for language model inference, marking OpenAI's first venture into custom hardware design.
The chip was designed from scratch in nine months and will be deployed at gigawatt scale by late 2026, reflecting OpenAI's strategy to control its full technology stack from chip to product for faster, more reliable, and lower-cost operations.
What happened
OpenAI and Broadcom unveiled Jalapeño, OpenAI's first custom chip built from scratch for language model inference. The two companies are building a multi-generation platform together, with Broadcom handling manufacturing and networking, and Celestica managing boards and system integration. The design cycle took nine months, which OpenAI says is the fastest ASIC development cycle for high-performance semiconductors it is aware of.
Why it matters
OpenAI argues that controlling the full technology stack from chip to product allows it to run models faster, more reliably, and at lower cost. This move signals OpenAI's shift from focusing only on models and products into custom hardware—a strategy that may let it reduce dependence on existing chip suppliers and potentially improve the economics of running its AI services at scale.
What to watch
Early tests showed performance per watt that OpenAI claims is "substantially better" than current state-of-the-art hardware, though these are self-reported numbers that have not been independently verified and a technical report is expected to follow. The first deployment is planned for late 2026 at gigawatt scale, together with Microsoft and other partners.
Ask the AI about this article →
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Nvidia's stock is trading about 13.5% below its estimated intrinsic value of $251 per share, based on a Discou…

Elon Musk warned that roughly 15 gigawatts of AI compute scheduled for next year may never get powered up in 2…

JetBrains announced Junie Local, a coding agent that runs entirely on a local machine

A new poll from Embold Research and Heatmap Pro shows 75% of Americans now oppose building AI data centers in…

A Pew Research Center survey of 3,488 U.S

Caterpillar is applying lessons from its mining automation business to broader AI deployment, including a voic…
