the.bay.news

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

TechCrunch
OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show
Tested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.

0 comments

Sign in to join the discussion — your thebay.events account works here.

No comments yet.