{
  "id": 567493,
  "title": "Nvidia launches a smaller, faster Nemotron model and a router to put it to work",
  "url": "https://urgent.news/2026/08/11/nvidia-launches-a-smaller-faster-nemotron-model-and-a-router-to-put",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-11T13:00:00.000Z",
  "source": {
    "name": "The New Stack",
    "slug": "the-new-stack",
    "url": "https://thenewstack.io/nvidia-nemotron-lightning-switchyard/"
  },
  "original_language": "en",
  "account": "Nvidia has introduced Nemotron 3.5 Lightning, a 30-billion-parameter mixture-of-experts model, and NeMo Switchyard, an open-source library for routing models. Nemotron 3.5 Lightning, developed with contributions from the Nemotron coalition, exhibits reasoning capabilities similar to the larger Nemotron 3 Super model, but falls short of Google's Gemma 4 31B model in terms of performance. Nvidia emphasizes speed and customization, claiming that 3.5 Lightning can deliver up to 4x faster output speeds and can be finely tuned for specific workflows. Post-training customization significantly improves accuracy and outperforms proprietary models in specialized tasks. NeMo Switchyard, a new routing library, allows developers to define model pools and set routing criteria based on quality, latency, or cost. In Nvidia's internal benchmarks, a Switchyard-routed system using a combination of open models and Anthropic's Opus 4.8 maintained frontier-level accuracy while cutting task-completion costs to about a third of running Opus alone. The Nemotron 3.5 Lightning model is available on Hugging Face, ModelScope, OpenRouter, and Nvidia's NIM microservice platform, while NeMo Switchyard is open-source and available on GitHub, with more partner integrations planned.",
  "summary": "Nvidia on Tuesday launched Nemotron 3.5 Lightning, the newest member of its Nemotron 3 family of open models. In addition, The post Nvidia launches a smaller, faster Nemotron model and a router to put it to work appeared first on The New Stack .",
  "key_points": [
    "Nvidia releases Nemotron 3.5 Lightning, a 30-billion-parameter model.",
    "NeMo Switchyard, an open-source routing library, enables model optimization.",
    "Combined system achieves frontier-level accuracy at reduced costs."
  ],
  "editors_take": "Nvidia's new Nemotron model and NeMo Switchyard library enable faster and more customized AI workflows, allowing developers to cut costs while maintaining accuracy, and giving them more control over model deployment.",
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}