OpenAI posts first Jalapeño measurements: more work per watt than GB200/GB300
At Hot Chips 2026 OpenAI shared first results from its custom inference ASIC with Broadcom. SemiAnalysis InferenceX on GPT-OSS 120B, DeepSeek R1, and Kimi K2.5 1T showed 1.5–1.9× more AI work per watt at peak throughput.



