OpenAI Jalapeño: Better than Nvidia Blackwell
7.4 relevance
Score Breakdown
technical depth 8
novelty 9
actionability 3
community 9
strategic 8
personal 9
Scored daily by a customisable AI persona to surface the most relevant engineering leadership news.
OpenAI's custom inference chip directly impacts AI agent cost and performance, a core interest.
Summary
OpenAI's custom inference ASIC 'Jalapeño,' built with Broadcom in just 16 months, beats Nvidia Blackwell and AMD chips on throughput per watt across all tested open-source models, including DeepSeek R1 and Kimi-K2. The chip uses HBM4, achieves over 700 tokens/sec/user at concurrency 1 with single-token prediction, and is a general-purpose inference accelerator—not specialized for OpenAI's models. Performance was verified in-lab using the InferenceX benchmark suite, though AgentX results are pending.
Author
Bryan Shan — HPC