{
  "id": 11720412,
  "title": "Jev as a Tool Router: Cutting Agent Cost Without Killing the Investigation",
  "url": "https://urgent.news/2026/10/03/jev-as-a-tool-router-cutting-agent-cost-without-killing-the",
  "topic": "ai",
  "section": "AI",
  "published": "2026-10-03T16:13:59.000Z",
  "source": {
    "name": "Dev.to",
    "slug": "dev-to",
    "url": "https://dev.to/karthikbommineni/jev-as-a-tool-router-cutting-agent-cost-without-killing-the-investigation-2ll8"
  },
  "original_language": "en",
  "account": "In this experiment, the author sought to determine if utilizing the System 1 model Jev could help reduce costs and maintain investigation quality when dealing with large tool catalogs in AI agents. A comparison was made between using Kimi K3 as the primary decision-making model with the full tool catalog, Jev routing the tool selection decision, and the Astra model using the full catalog.\n\nThe experiment involved three scenarios, each with three catalog sizes: 50, 100, and 200 tools. Each scenario had the same task and consistent tool results, with the final outcome being scored based on the incident note and the actions taken. For the 200 tool catalog, Jev was forced to make a single tool selection per round, while full-menu LLMs could make multiple selections in a single hop.\n\nThe results showed that Jev + Kimi maintained strict golden pass scores at all catalog sizes, while the Astra full-menu option only achieved this at 50 tools and failed the remaining tests. Moreover, the cost for the Jev + Kimi setup was significantly lower than Astra at all catalog sizes. For the 200 tool catalog, Astra cost $0.228 compared to $0.0034 for Jev + Kimi. Additionally, the number of hops required by Jev was consistently lower than the Astra full-menu option, indicating better efficiency in the routing process.",
  "summary": "By now, you have probably heard about Jev, a System 1 model that has been getting a lot of attention lately. In simple terms, a System 1 model is built for fast, cheap, bounded decisions, while a System 2 model is the slower, heavier LLM that reasons through open-ended work (the big 3: ChatGPT, Claude, Gemini (or even Grok)). I will not dig into Jev's architecture in this post; that is probably a…",
  "key_points": [
    "Jev + Kimi maintained golden pass scores across all catalog sizes",
    "Jev + Kimi setup cost significantly lower than Astra at all sizes",
    "Jev required fewer hops than Astra full-menu option for tool selection"
  ],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}