Observed Signal · Jun 24, 2026 · Product Launch · Source: OpenAI Blog · Impact: 4/5 · Sentiment: Positive
OpenAI and Broadcom unveil Jalapeño inference chip
OpenAI and Broadcom announced Jalapeño, an LLM-optimized accelerator described as OpenAI’s first 'Intelligence Processor' and the initial component of a multi-generation compute platform co-developed with Broadcom and Celestica. OpenAI says the chip was designed from the ground up for modern and future LLM inference, produced from design to tape-out in nine months with assistance from OpenAI models, and that engineering samples are running ML workloads (including GPT‑5.3‑Codex‑Spark). Early internal testing reportedly shows performance per watt substantially better than current state-of-the-art; a detailed technical report will follow. The partners plan gigawatt-scale deployments with data center partners (including Microsoft) beginning in 2026, aiming to lower inference cost, latency, and improve reliability for large-scale interactive LLM products.
Major platform (OpenAI) announced a purpose-built LLM inference accelerator co-developed with Broadcom that claims materially better performance-per-watt and plans gigawatt-scale deployments; this can reduce inference costs, change infrastructure sourcing, and affect the economics and availability of AI-powered products across industries.
Track OpenAI Signals & Market Shifts in Real-Time
Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.
Key Takeaways & Evidence Grounding
- OpenAI and Broadcom unveiled Jalapeño, described as OpenAI’s first 'Intelligence Processor' accelerator optimized for LLM inference.
- Jalapeño was co-developed with Broadcom and Celestica and moved from initial design to manufacturing tape-out in nine months.
- Engineering samples of Jalapeño are running ML workloads in-lab at target frequency and power, including GPT‑5.3‑Codex‑Spark.
- OpenAI says early testing shows Jalapeño will deliver performance per watt substantially better than current state-of-the-art; a detailed technical report is forthcoming.
- The partners plan multi-generation, gigawatt-scale deployments with data center partners (including Microsoft) beginning in 2026.
Connected Companies & Entities
3 Entities mapped“OpenAI and Broadcom (NASDAQ: AVGO) today unveiled Jalapeño, OpenAI’s first Intelligence Processor: an accelerator architected around OpenAI’...”
“OpenAI and Broadcom (NASDAQ: AVGO) today unveiled Jalapeño, OpenAI’s first Intelligence Processor: an accelerator architected around OpenAI’...”
“By co-developing our industry-leading silicon directly with OpenAI, we are enabling the deployment of gigawatt scale data centers with Micro...”
Ontology Mapping & Concepts
Related Market Signals & Shifts
Recent verified developments and strategic activity across this market segment.
OpenAI unveils Jalapeño inference chip with Broadcom
OpenAI unveiled its first custom-built inference processor, called Jalapeño, developed in collaboration with Broadcom. The chip is designed specifically for inference workloads and, according to OpenAI, early tests show substantially better performance-per-watt than current alternatives. OpenAI said its own AI models assisted chip development. The partnership with Broadcom was announced previously in October, and the move is widely seen as a way for OpenAI to reduce reliance on Nvidia GPUs for inference, while heavier tasks like pre‑training will likely continue to use existing GPU hardware. OpenAI framed the chip as part of a broader strategy to optimize across the stack — from chip architecture to deployment systems — to make models faster, more reliable, and cheaper to run.
OpenAI and Broadcom Reveal Jalapeño AI Chip
OpenAI and Broadcom publicly unveiled Jalapeño, their first jointly developed custom AI inference accelerator, positioning it as an “Intelligence Processor” designed to improve performance‑per‑watt and lower token costs for large language model inference. Broadcom has already delivered engineering silicon to OpenAI; Broadcom CEO Hock Tan expects initial deployment late 2026, a ramp in 2027 and full-scale deployment in 2028. The announcement arrives as Broadcom shares have fallen sharply since early June, and follows market reports that hyperscalers (Alphabet, ByteDance) are exploring additional chip design partners to diversify supply. OpenAI’s Greg Brockman emphasized Jalapeño is complementary to, not a replacement for, GPU-based solutions from firms like Nvidia. The collaboration signals increased vertical integration by an AI provider and a potential pathway to reduce inference costs and diversify compute beyond dominant GPU suppliers.
OpenAI’s Jalapeño Chip Shows Leading Inference Efficiency
OpenAI revealed Jalapeño — its first custom inference ASIC and rack-scale platform co-developed with Broadcom and Celestica and shown at Hot Chips — with A0 engineering samples taped out Nov 2025. The compute die uses TSMC N3P, MXFP numeric formats and HBM4 (15.4 TB/s per package); package TDP is ~700 W with sustained test power ≲ ~550 W. The rack design keeps model state (KV cache) local, simplifies on-node fabric and can scale to 2,048 XPUs. OpenAI and SemiAnalysis/InferenceX benchmarks report substantial performance-per-watt and latency gains (and partial advantages vs Nvidia GB300), but results are not independently verified and did not include Nvidia Vera Rubin. OpenAI targets small-volume deployment end‑2026 and broader ramp in 2027, pursuing a multi‑vendor production strategy while production economics, yield and fleet reliability remain unproven.
Track Real-Time Market Signals & Shifts
Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.
