OpenAI Unveils Jalapeño AI Chip Benchmarks High Throughput with Ultra-Low Latency.
OpenAI Reveals Test Results for Custom 'Jalapeño' AI Chip Developed with Broadcom and Celestica OpenAI has officially disclosed performance benchmarks for Jalapeño , its custom artificial intelligence chip co-designed in partnership with Broadcom and Celestica . Initially unveiled in June without granular technical specifications, new data showcases significant operational efficiency gains in handling large-scale inference workloads. According to Richard Ho , Vice President of Hardware at OpenAI, the defining architectural advantage of Jalapeño is its ability to break the traditional engineering trade-off between throughput (processing volume per unit) and latency (response speed). Typically, AI chips must sacrifice execution speed to maximize processing capacity, or vice versa. To evaluate real-world performance, OpenAI utilized the InferenceX benchmark suite across three open-source AI models: GPT-OSS 120B , DeepSeek R1 , and Kimi K2.5 1T . Benchmark results demonstrated...