OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show
TechCrunch AIen

Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.
This is a short summary published by AI Global Wire. The full article is owned and hosted by TechCrunch AI — open it there to read it in full.
Read the full story at TechCrunch AI- OpenAI
Related AI news
- A detailed look at Jalapeño, OpenAI's ASIC developed with Broadcom in 16 months, which beat Nvidia, AMD, and Google chips on multiple top open-source models (SemiAnalysis)Techmeme · August 25, 2026
- OpenAI says its Jalapeño chip delivered 1.5x-1.9x more AI work per watt and 1.7x-3.6x lower latency vs. Nvidia chips across GPT-OSS, DeepSeek R1, Kimi K2.5 1T (Emma Roth/The Verge)Techmeme · August 25, 2026
- OpenAI says its Jalapeño chip can power faster AI responses than the competitionThe Verge AI · August 25, 2026
- OpenAI bans a cluster of Russian ChatGPT accounts that used VPNs to evade restrictions and run an influence operation, including creating social media comments (Kai Nicol-Schwarz/CNBC)Techmeme · August 25, 2026
- ‘The world seems to be ready’: An interview with OpenAI head of product Thibault SottiauxTechCrunch AI · August 25, 2026
- Apple nimmt zu OpenAIs Versuchen Stellung, seine Klage abzuweisenheise online – KI · August 25, 2026