OpenAI and Broadcom announced Jalapeño on Wednesday, a custom application-specific integrated circuit designed for large language model inference at data-center scale. The companies said development took nine months and incorporated OpenAI researchers' projections for future model architectures and serving patterns.
Jalapeño marks OpenAI's first disclosed processor partnership, joining a wave of frontier labs seeking dedicated silicon as cloud GPU supply tightens. Broadcom, already a major supplier of custom chips to hyperscalers, will manufacture the ASIC for deployment in OpenAI-aligned facilities by year end, both companies said.
Technical Claims
OpenAI said early testing shows Jalapeño delivering substantially better performance per watt than current state-of-the-art inference hardware, though it has not finished benchmarking and plans a detailed technical report in coming months. Broadcom emphasized the design was built from scratch for transformer inference rather than adapted from general matrix-multiplication accelerators.
Engineers familiar with custom AI silicon said nine-month cycles are aggressive; success depends on whether OpenAI locked specifications early enough to avoid respins when model context lengths grow. The chip focuses on inference, not training — the phase that consumes most Nvidia HBM-equipped GPUs today.
Why Custom Silicon Now
Hyperscalers and model labs face a compute crunch as agentic applications multiply token volumes per user task. Google, Amazon, and Microsoft already deploy in-house inference chips alongside Nvidia clusters. OpenAI's Broadcom partnership signals it will not rely exclusively on rented GPU capacity as ChatGPT, Codex, and enterprise API usage expand.
Capital markets reacted modestly: Broadcom shares rose 2.1 percent while Nvidia slipped 0.8 percent on fears of long-term inference share erosion. Analysts noted Jalapeño volumes will remain small relative to Nvidia data-center revenue through 2027.
Supply-Chain Angle
The announcement arrives as memory shortages constrain consumer electronics and as U.S. export rules reshape who may access frontier software. Custom inference chips reduce dependence on HBM-heavy GPU configurations for serving workloads, potentially easing one bottleneck if designs reach volume production.
Broadcom did not disclose process node or packaging partners. Taiwan Semiconductor Manufacturing Company supplies most advanced Broadcom products; any production ramp would compete for fab attention with Apple and Nvidia orders.
Deployment Timeline
Both companies said Jalapeño will appear in data centers before December, without naming sites. OpenAI operates and leases capacity across the United States and partners with Microsoft Azure for overflow. Operators will watch whether Jalapeño clusters reduce per-token serving cost enough to change API pricing.
For the broader industry, Jalapeño validates inference-specific ASICs as a complement — not yet a replacement — for GPU fleets built over the past three years of generative AI expansion.



