25 ago OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show
Posted at 11:22h
in Tech-En
Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.