Tech & AI News
TechMeme

OpenAI says its Jalapeño chip delivered 1.5x-1.9x more AI work per watt and 1.7x-3.6x lower latency vs. Nvidia chips across GPT-OSS, DeepSeek R1, Kimi K2.5 1T (Emma Roth/The Verge)

OpenAI reports that its Jalapeño AI inference chip achieves 1.5-to-1.9 times higher work-per-watt efficiency and 1.7-to-3.6 times lower latency compared with Nvidia’s competing GPUs when running GPT-OSS, DeepSeek R1, and Kimi K2.5 1T models. The benchmark results highlight a substantial performance advantage for OpenAI’s custom silicon in power-constrained and latency-sensitive workloads.