{
  "id": 3281215,
  "title": "OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show",
  "url": "https://urgent.news/2026/08/25/openais-jalapeno-chip-is-built-for-fast-inference-at-scale-benchmarks",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-25T14:22:04.000Z",
  "source": {
    "name": "TechCrunch",
    "slug": "techcrunch",
    "url": "https://techcrunch.com/2026/08/25/openais-jalapeno-chip-is-built-for-fast-inference-at-scale-benchmarks-show/"
  },
  "original_language": "en",
  "account": "At the Hot Chips conference on Tuesday, OpenAI unveiled Jalapeño, a new chip designed for rapid inference at scale. The company released initial benchmark results for the chip, which outperformed the current state-of-the-art inference processors in terms of tokens per user and throughput per kilowatt. Richard Ho, OpenAI's head of hardware, stated, \"The bottom line is that the results show a very, very significant performance advance over state of the art.\" The benchmark tests were conducted against an Nvidia Blackwell system, but by the time Jalapeño is fully deployed, the competitive landscape may have shifted. OpenAI estimated that Jalapeño would enter limited production by the end of 2026, with wider release in 2027. Developed jointly with Broadcom, OpenAI's own models played a role in the chip's development. The company aims to create a multigenerational platform, integrating AI products, models, chips, and memory. This full-stack approach allowed OpenAI to tackle specific challenges in the inference process, particularly reducing delays in the prefill and communication phases. By minimizing data movement and communication delays, Jalapeño can keep model state, like the KV cache used during response generation, local while activating the optimal compute, memory, and networking resources for each inference phase.",
  "summary": "Tested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 5,
    "also_reported_by": [
      {
        "outlet": "The Register",
        "title": "OpenAI's upcoming Jalapeño chip looks like it'll be an inference beast",
        "url": "https://urgent.news/2026/08/25/openais-upcoming-jalapeno-chip-looks-like-itll-be-an-inference-beast",
        "published": "2026-08-25T14:00:00.000Z"
      },
      {
        "outlet": "The Verge",
        "title": "OpenAI says its Jalapeño chip can power faster AI responses than the competition",
        "url": "https://urgent.news/2026/08/25/openai-says-its-jalapeno-chip-can-power-faster-ai-responses-than-the",
        "published": "2026-08-25T14:00:00.000Z"
      },
      {
        "outlet": "Techmeme",
        "title": "A detailed look at Jalapeño, OpenAI's ASIC developed with Broadcom in 16 months, which beat Nvidia, AMD, and Google chips on multiple top open-source models (SemiAnalysis)",
        "url": "https://urgent.news/2026/08/25/a-detailed-look-at-jalapeno-openais-asic-developed-with-broadcom-in",
        "published": "2026-08-25T15:15:02.000Z"
      },
      {
        "outlet": "Digital Trends",
        "title": "OpenAI says its Jalapeno AI chip delivers faster responses than rivals like Nvidia",
        "url": "https://urgent.news/2026/08/25/openai-says-its-jalapeno-ai-chip-delivers-faster-responses-than",
        "published": "2026-08-25T15:24:59.000Z"
      }
    ]
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}