OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

发生了什么
Tested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.
摘要按规则整理自下方来源原文