{
  "id": 3277889,
  "title": "OpenAI's upcoming Jalapeño chip looks like it'll be an inference beast",
  "url": "https://urgent.news/2026/08/25/openais-upcoming-jalapeno-chip-looks-like-itll-be-an-inference-beast-3277889",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-25T14:00:00.000Z",
  "source": {
    "name": "The Register Science",
    "slug": "the-register-science",
    "url": "https://www.theregister.com/systems/2026/08/25/openais-upcoming-jalapeno-chip-looks-like-itll-be-an-inference-beast/5292052"
  },
  "original_language": "en",
  "account": "OpenAI unveiled its upcoming Jalapeño AI accelerator chip at the Hot Chips semiconductor development conference in Stanford. This chip, developed in collaboration with Broadcom, is the first of a series of custom silicon from OpenAI designed for AI inference tasks. Compared to Nvidia's GPU systems, OpenAI claims Jalapeño will deliver higher throughput and lower latency when it becomes available later this year and reaches volume production in 2027.\n\nThough Jalapeño won't replace OpenAI's existing hardware partners, it's designed with AI in mind. Memory bandwidth is crucial for inference, and Jalapeño appears to excel in this area. Early benchmarks show the chip delivering 1.5x to 1.9x more AI work at peak throughput and 1.7x to 3.6x lower end-to-end latency than competitors like GPT-OSS-120B, DeepSeek R1, and Kimi K2.5. For ultra-low-latency inference, Jalapeño is 2.1x to 4.1x faster.\n\nEach Jalapeño system contains 128 accelerators, providing 1.7 exaFLOPS of 4-bit compute, 27.5 TB of HBM4 memory, and nearly 2 petabytes per second of memory bandwidth. Compared to AMD and Nvidia's latest rack systems, Jalapeño uses between 40 and 60 percent of the power. The system design is based on a rack scale architecture similar to Nvidia's NVL72 or AMD Helios.",
  "summary": "128 chips, 1.7 exaFLOPS, and 27 TB of HBM give Altman and crew a leg up over Blackwell, and maybe even Rubin",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 6,
    "also_reported_by": [
      {
        "outlet": "The Register",
        "title": "OpenAI's upcoming Jalapeño chip looks like it'll be an inference beast",
        "url": "https://urgent.news/2026/08/25/openais-upcoming-jalapeno-chip-looks-like-itll-be-an-inference-beast",
        "published": "2026-08-25T14:00:00.000Z"
      },
      {
        "outlet": "The Verge",
        "title": "OpenAI says its Jalapeño chip can power faster AI responses than the competition",
        "url": "https://urgent.news/2026/08/25/openai-says-its-jalapeno-chip-can-power-faster-ai-responses-than-the",
        "published": "2026-08-25T14:00:00.000Z"
      },
      {
        "outlet": "TechCrunch",
        "title": "OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show",
        "url": "https://urgent.news/2026/08/25/openais-jalapeno-chip-is-built-for-fast-inference-at-scale-benchmarks",
        "published": "2026-08-25T14:22:04.000Z"
      },
      {
        "outlet": "Techmeme",
        "title": "A detailed look at Jalapeño, OpenAI's ASIC developed with Broadcom in 16 months, which beat Nvidia, AMD, and Google chips on multiple top open-source models (SemiAnalysis)",
        "url": "https://urgent.news/2026/08/25/a-detailed-look-at-jalapeno-openais-asic-developed-with-broadcom-in",
        "published": "2026-08-25T15:15:02.000Z"
      },
      {
        "outlet": "Digital Trends",
        "title": "OpenAI says its Jalapeno AI chip delivers faster responses than rivals like Nvidia",
        "url": "https://urgent.news/2026/08/25/openai-says-its-jalapeno-ai-chip-delivers-faster-responses-than",
        "published": "2026-08-25T15:24:59.000Z"
      }
    ]
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}