OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show
Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.
Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art. This story matters for AI & Agent Economy readers tracking ip. Reported by techcrunch.com. Read the full original at the source link below.
Originally reported by techcrunch.com. IPNews curates and briefs the ai & agent economy stories that matter. Our editorial policy →