OpenAI Jalapeño Chip Crushes Nvidia Blackwell: 1.9x Throughput & 3.6x Latency Drop in First Benchmarks
OpenAI's first custom inference chip delivers 1.5-1.9x more AI work per watt and 1.7-3.6x lower latency than Nvidia's flagship Blackwell GPU across GPT-OSS 120B, DeepSeek R1, and Kimi K2.5 1T models.