OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show
Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.
本文内容来源于互联网,版权归原作者所有
查看原文