All NewsEducationTV
Equities & FundsCrypto & Digital AssetsAI & TechnologyBusiness & CorporateUS Politics & PolicyGeopolitics & Global RiskMacro, Rates & FXCommodities & EnergyEuropean Politics & MarketsAsia-PacificReal Estate & Property
All NewsHome
← Back to AI & Technology

OpenAI's Jalapeño chip benchmarks show performance gains over Nvidia Blackwell

Created at 25 Aug · 2:56 PM1 source↑ Market-relevant
IN SHORT

OpenAI has released benchmark results for its custom Jalapeño chip, indicating significant performance advantages in tokens per user and throughput per kilowatt compared to Nvidia's Blackwell inference processors. The chip, developed with Broadcom, is slated for limited deployment in late 2026.

Key Numbers

2026Jalapeño deployment year
2027Significant Jalapeño deployment year

Who's Involved

OpenAI
Developer of the Jalapeño inference chip
Richard Ho
Head of hardware at OpenAI
Broadcom
Collaborator in Jalapeño chip development
Nvidia
Manufacturer of Blackwell inference processors
OpenAI's Jalapeño chip benchmarks show performance gains over Nvidia Blackwell

↳ Why This Matters

The development of specialized AI inference chips like OpenAI's Jalapeño signifies a move towards more efficient and powerful AI hardware, potentially reducing costs and improving user experience for AI applications.

Key facts

  • OpenAI's Jalapeño chip shows performance advantages over Nvidia's Blackwell inference processors in benchmarks.
  • Jalapeño offers more tokens per user and higher throughput per kilowatt.
  • The chip is designed to minimize delays in prefill and communication phases of AI inference.
  • Jalapeño was developed in collaboration with Broadcom.
  • Limited deployment is anticipated for late 2026, with significant deployment in 2027.
  • OpenAI has revealed new benchmark results for its custom-designed Jalapeño chip, showcasing significant performance gains in artificial intelligence inference processing. At the Hot Chips conference, OpenAI presented data indicating that Jalapeño surpasses current state-of-the-art inference processors in both tokens per user and throughput per kilowatt. Richard Ho, OpenAI's head of hardware, described the performance advance as "very, very significant," highlighting the chip's efficiency in serving multiple customers with low latency.

    The benchmarks compare Jalapeño against Nvidia's Blackwell system. OpenAI developed Jalapeño in collaboration with Broadcom, utilizing its own AI models to aid in the chip's development. The company plans for Jalapeño to be a multigenerational platform, integrating AI products, models, chips, and memory development. A key design focus for Jalapeño is to reduce delays in the prefill and communication phases of inference, often bottlenecks in processing. This is achieved by minimizing data movement and keeping model state local to optimize compute, memory, and networking for each inference stage.

    OpenAI estimates that Jalapeño will see limited deployment at the end of 2026, with more substantial rollout expected in 2027.

    Frequently asked questions

    Jalapeño is a custom-designed chip developed by OpenAI, in collaboration with Broadcom, specifically for accelerating AI inference at scale.

    Benchmarks show Jalapeño offers more tokens per user and higher throughput per kilowatt than current state-of-the-art inference processors, including Nvidia's Blackwell system.

    OpenAI anticipates limited deployment by the end of 2026, with more significant deployment expected in 2027.

    Jalapeño is designed to minimize data movement and communication delays, particularly in the prefill and communication phases of inference, by keeping model state local.

    What Happens Next

    01Limited deployment of Jalapeño is expected by the end of 2026.
    02More significant deployment of Jalapeño is anticipated in 2027.

    How It Developed

    OpenAI shared detailed information and benchmark results for its Jalapeño chip at the Hot Chips conference.
    Jalapeño demonstrated superior tokens per user and throughput per kilowatt compared to state-of-the-art inference processors.
    The chip's performance was benchmarked against an Nvidia Blackwell system.
    Jalapeño is designed to minimize data movement and communication delays during inference.
    OpenAI collaborated with Broadcom on Jalapeño's development, with OpenAI's models assisting in the process.
    Limited deployment of Jalapeño is expected by the end of 2026, with broader deployment in 2027.

    Sources

    T1
    OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks showTechCrunch

    Related Stories

    Apple unveils Mac Studio, Mac mini with M5 Ultra and M6 chips for AI
    25 Aug · 1:12 PM
    AI data centers drive demand for new solid-state transformer technology
    24 Aug · 9:36 PM
    AI Founders Launch Physics Model After Rejecting Bezos's Project Prometheus
    25 Aug · 10:05 AM
    OpenAI aims to bring AI agents to all workers with ChatGPT Work
    25 Aug · 12:16 PM
    DeepSeek leads surge in low-cost Chinese open-weight models on US platform
    25 Aug · 10:36 AM