OpenAI's Jalapeño has benchmarks now. The watt story is the real one.
OpenAI shared first benchmark results for Jalapeño at Hot Chips 2026: 1.5–1.9x more throughput per watt and 1.7–3.6x lower latency than Nvidia GB300. The chip runs at 700W versus roughly 1,400W for Nvidia's flagship. Mass deployment comes in 2027, but the direction of API pricing is already set.