Jalapeño: OpenAI's first custom inference chip, built with Broadcom
freshWhat: OpenAI and Broadcom unveiled "Jalapeño," an Intelligence Processor designed from scratch for LLM inference. Nine-month tape-out, accelerated by OpenAI's own models. Engineering samples running GPT-5.3-Codex-Spark at production target frequency. TSMC manufacturing, Celestica on boards/racks. Deployment planned at gigawatt scale late 2026.
Why it matters: OpenAI is vertically integrating the full stack — chip to model to product. Broadcom CEO Hock Tan claims Jalapeño is "just as good" as NVIDIA Blackwell for inference, though no independent benchmarks exist yet. Microsoft reportedly committed to buying 40% of first-run chips. This directly challenges NVIDIA's inference dominance and changes OpenAI's cost structure permanently.
Watch: Whether independent performance data validates the per-watt claims; whether the specialist (inference-only) architecture limits flexibility if model architectures shift.