FireTofu
OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

Technology · en

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

TechCrunch · Aug 25, 2026, 2:22 PM UTC

Tested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.