techcrunch.com4 weeks agoOpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks showTested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.Visit techcrunch.comBookmarkAdd to collection