Skip to content

OpenAI Jalapeño: Better than Nvidia Blackwell

7.4 relevance
Score Breakdown
technical depth
8
novelty
9
actionability
3
community
9
strategic
8
personal
9

Scored daily by a customisable AI persona to surface the most relevant engineering leadership news.

OpenAI's custom inference chip directly impacts AI agent cost and performance, a core interest.

AI/ML newsletter.semianalysis.com
OpenAI Jalapeño: Better than Nvidia Blackwell
Summary

OpenAI's custom inference ASIC 'Jalapeño,' built with Broadcom in just 16 months, beats Nvidia Blackwell and AMD chips on throughput per watt across all tested open-source models, including DeepSeek R1 and Kimi-K2. The chip uses HBM4, achieves over 700 tokens/sec/user at concurrency 1 with single-token prediction, and is a general-purpose inference accelerator—not specialized for OpenAI's models. Performance was verified in-lab using the InferenceX benchmark suite, though AgentX results are pending.

Author

Bryan Shan — HPC