OpenAIβs JalapeΓ±o chip is built for fast inference at scale, benchmarks show
Tested on Semianalysisβs InferenceX benchmark, JalapeΓ±o registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.
Log in to bookmark articles and create collections
Isabella News