{
  "id": 3544604,
  "title": "Hot Chips 2026: Nvidia presents Groq 3 LPX architecture and unveils its first third-party inference benchmark — LP30-based rack already in production, company says",
  "url": "https://urgent.news/2026/08/26/hot-chips-2026-nvidia-presents-groq-3-lpx-architecture-and-unveils",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-26T16:23:37.000Z",
  "source": {
    "name": "Tom's Hardware",
    "slug": "tom-s-hardware",
    "url": "https://www.tomshardware.com/tech-industry/semiconductors/nvidia-presents-groq-3-lpx-architecture-and-unveils-its-first-third-party-inference-benchmark"
  },
  "original_language": "en",
  "account": "At the Hot Chips 2026 conference, Nvidia unveiled its first third-party inference benchmark for their newly acquired Groq 3 LPX architecture. Former Groq chief architect Igor Arsovski, now Nvidia's VP of hardware, presented the architecture and benchmark results. The Groq 3 LPX rack achieved 3,431 output tokens per second on a 100K-context Gemma 4 31B reasoning workload, outperforming the next-fastest public endpoint by four times. The rack, built on Nvidia's LP30 chip obtained through the $20 billion Groq deal, is already in production. The LP30 chip carries 500MB of on-die SRAM and no HBM, leading to a fully deterministic pipeline and higher token generation rates. Nvidia claims a 10 to 11% performance boost under the same thermal limit due to deterministic execution and heat equalization. The LPX rack is designed to work with Nvidia's Vera Rubin NVL72 system, with Rubin GPUs handling compute-heavy prefill phases and KV cache while LPU generates output tokens.",
  "summary": "Igor Arsovski, now Nvidia's VP of hardware, presented the Groq 3 LPX rack's architecture and published the first third-party benchmark of the hardware.",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 3,
    "also_reported_by": [
      {
        "outlet": "Techmeme",
        "title": "Indian AI infrastructure company AM Intelligence orders 9,000 Nvidia Vera Rubin systems and plans to offer 1GW of computing capacity as part of an $8B project (Saritha Rai/Bloomberg)",
        "url": "https://urgent.news/2026/08/26/indian-ai-infrastructure-company-am-intelligence-orders-9-000-nvidia",
        "published": "2026-08-26T05:25:02.000Z"
      },
      {
        "outlet": "Free Press Journal",
        "title": "Indian Firm AM Intelligence Places Order For 9,000 Nvidia Vera Rubin System Chips For Hyderabad AI Factory",
        "url": "https://urgent.news/2026/08/26/indian-firm-am-intelligence-places-order-for-9-000-nvidia-vera-rubin",
        "published": "2026-08-26T11:26:59.000Z"
      }
    ]
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}