FNU FastNewsUpdateHub

Tech, devices and platforms — updated fast

TechCrunch

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

Tested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.

Photo: TechCrunch

Tested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.

Originally published by TechCrunch. Read the full story at the source →

More stories