TechCrunch
TechCrunch
US · 27 mins ago

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

Tested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.
TechCrunch
Do you trust TechCrunch?
Sign in to rate
Discussion
?

No comments yet — be the first to start the discussion!