Full Analysis: OpenAI Jalapeño chip..
OpenAI's custom inference chip, Jalapeno, designed with Broadcom in 13 months, shows preliminary benchmarks that beat Nvidia Blackwell on specific inference tasks like serving Kim K 2.5 at 100 tokens per second, with nearly nine times the throughput. However, the chip is an ASIC optimized for inference, not general-purpose compute, and its advantages are narrower than they appear: it trails Blackwell in raw FP4 compute, and OpenAI has only released results for one benchmark (8K1K), leaving questions about agentic workloads and broader stability. The real story is that OpenAI is betting on specialized, power-efficient inference hardware to address energy and cost bottlenecks, but it's too early to declare Nvidia dethroned.