Beta The Briev beta is out. Free on iPhone via TestFlight — install it in under a minute.

Join the beta ↗
Briev
Live
Technology

OpenAI unveils Jalapeño ASIC promising faster, more efficient AI inference

OpenAI announced its new Jalapeño ASIC, claiming it delivers lower latency and higher throughput than competing AI chips.

In a blog post and subsequent briefing, OpenAI introduced its Jalapeño application-specific integrated circuit, designed for AI inference tasks. Partnered with Broadcom, the chip was measured on the InferenceX platform and beat Nvidia's top-tier GB200/GB300 superchips, delivering between 1.5 and 1.9 times more AI work per watt on models such as GPT-OSS 120B, DeepSeek R1 and Kimi K2.5. Latency improvements ranged from 1.7 to 3.6 times lower, which the company says will enable quicker responses and more reliable access as demand rises.

OpenAI intends to ship Jalapeño in small volumes before the end of the year and scale up production through 2027, while still relying on partners like Nvidia for other compute needs. Development of second- and third-generation versions is already underway.

Why it matters

Faster, more efficient AI chips can lower costs and improve responsiveness for users of AI services.

In this story

OpenAIJalapeño chipAI inferencelatencythroughputBroadcom partnershipNvidia superchipsbenchmarkAI workload efficiency
Get the beta ↗