OpenAI’s custom inference chip, Jalapeño, delivered significantly higher throughput and lower latency than Nvidia’s Blackwell system in early benchmarks, according to data presented.