OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

TechCrunch AIen

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.

This is a short summary published by AI Global Wire. The full article is owned and hosted by TechCrunch AI — open it there to read it in full.

Read the full story at TechCrunch AI
  • OpenAI

Related AI news