ChengRang

OpenAI Jalapeno

AI Platforms Paid

OpenAI first self-developed AI inference chip with Broadcom, TSMC 3nm, 9-month tape-out; ~50% lower inference cost than mainstream GPUs

OpenAIChipInferenceBroadcom3nm
Visit OpenAI Jalapeno

Disclaimer: Review content represents our editorial team's views and experience, not commercial recommendation or investment advice. Product info and pricing may change; refer to official sources.

Overview

Jalapeño is OpenAI's **first self-developed AI inference chip**, jointly announced with Broadcom on June 24, 2026. Manufactured using TSMC's 3nm process, it went from design to tape-out in just **9 months**—setting a new industry record. This marks the first time OpenAI has turned the rumor of a self-developed chip into reality among major tech companies, signaling a partial shift away from absolute reliance on NVIDIA GPUs.

Core positioning: **Specialized for LLM inference** (not training), targeting the "cost black hole" of large-scale API calls like ChatGPT. Broadcom CEO Hock Tan revealed at the launch: **In early lab tests, Jalapeño's inference cost is about 50% lower than mainstream GPUs, with performance comparable to NVIDIA Blackwell**—if this holds true in real-world deployment, it could halve the cost per ChatGPT call for OpenAI.

**Timeline and ecosystem**: Led by a former Google TPU veteran, co-designed with Broadcom, manufactured with assistance from Celestica, and taped out at TSMC. **Initial deployment is planned for late 2026**, and it is explicitly described as "**the first step in a multi-generation chip development plan**." Starting in 2026, OpenAI will collaborate with partners like Microsoft to drive **gigawatt-scale data center deployments**. OpenAI is still evaluating whether to sell the chip externally or keep it for internal use only.

Significance: **The emergence of Jalapeño marks the official transition of major AI companies from the 'buying chips' era to the 'making chips' era**—following Google TPU, Amazon Trainium, and Meta MTIA, OpenAI becomes another player joining the self-development camp, creating the first significant crack in NVIDIA's "shovel-selling" business.

Key Features

Use Cases

Pros

Pricing

**Not for external sale yet**: Currently only for OpenAI's internal use and deployment by deep partners like Microsoft. Whether to open sales externally is still under evaluation. **Indirect benefit**: Expected reduction in user costs for ChatGPT / API usage.

Summary

Jalapeño represents a critical step for OpenAI amid 'NVIDIA GPU shortages and soaring inference costs'—**9-month tape-out + 50% inference cost reduction + performance comparable to Blackwell**—moving AI giants from 'buying chips' to 'making chips'. Currently more of a strategic signal than a directly purchasable product; however, for ChatGPT users and OpenAI API developers, inference costs and availability may significantly improve within the next year. For domestic chip makers (Huawei Ascend, Cambricon, etc.), Jalapeño is a must-study benchmark.

Version History

Category
AI Platforms
Pricing
Paid
Tags
OpenAI · Chip · Inference
Website

Related Tools