OpenAI Unveils Jalapeño Custom Inference Chip and Expanded Compute Portfolio
Read the article: OpenAI Blog
OpenAI has published details about Jalapeño, its first custom inference chip, along with benchmark results showing the chip outperformed commercial alternatives on peak throughput per kilowatt and token latency when tested on GPT-OSS 120B via the public InferenceX benchmark. The chip also posted strong results on DeepSeek R1 and Kimi K2, suggesting the performance gains are not model-specific. OpenAI describes developing the chip, serving software, memory, and network as an integrated system, giving the company more direct control over the economics of running its models at scale.
The announcement also outlines OpenAI's broader compute portfolio, which now includes Microsoft, NVIDIA, AWS, AMD, Broadcom, Cerebras, CoreWeave, Oracle, SB Energy, and SoftBank. On the product side, GPT-5.6 Sol with max reasoning reached a new high score on the Artificial Analysis Coding Agent Index while using 54% fewer output tokens than a competing model. For legal professionals and enterprises relying on AI-powered tools for contract review, financial analysis, or document drafting, these efficiency gains translate directly into faster outputs, lower per-task costs, and more reliable performance on complex, multi-step workflows.