Urgent.News

What's breaking now, across thousands of outlets.

AI

Capsule Security fine-tunes Nvidia Nemotron models to stop rogue AI agents

Agentic artificial intelligence security startup Capsule Security Ltd. today released a detection system built on two Nvidia Corp. Nemotron models it fine-tuned itself, in what it calls an “AI circuit breaker” for rogue AI agents. The models judge an agent’s intended action in the moment before it executes. Customers can then allow it, flag it […] The post Capsule Security fine-tunes Nvidia…

Capsule Security fine-tunes Nvidia Nemotron models to stop rogue AI agents

Capsule Security, an agentic artificial intelligence security startup, has unveiled a detection system utilizing Nvidia Corp.'s Nemotron models, fine-tuned specifically for the purpose. This system, dubbed the "AI circuit breaker," operates by judging an agent's intended action in real-time before execution. Customers can subsequently allow, flag, or block the action in real-time, thereby establishing a control layer external to the agent.

This mechanism is designed to address the increasing prevalence of agents with access to sensitive data, source code, or production infrastructure. The system utilizes two of Nvidia's largest Nemotron models, Nemotron 3 Ultra, and Nemotron 3, which are capable of running within an agent's workflow without significantly hindering performance.

According to Capsule Security, the system achieved 98% accuracy on the StepShield benchmark, a metric for detecting rogue agent behavior at the step level. The benchmark utilized 9,429 code-agent trajectories derived from real incidents. While the startup acknowledges that rule-based guardrails can be effective, they often result in a high number of false positives.

In contrast, the fine-tuned models used by Capsule can classify a narrow set of actions with high precision and speed, making real-time decisions in as little as 71 milliseconds. The company claims that billions of tokens across millions of agent interactions are currently processed using this technology, with customers ranging from financial institutions to technology companies.

The startup's co-founders, Naor Paz and Lidan Hazout, emphasize that the core AI security risk now lies not just in what people can do with agents, but in what autonomous agents decide to do on their own. The release of this system follows the disclosure of two prompt injection vulnerabilities in widely used AI tools by Microsoft and Salesforce, both of which have since been patched.

Written by urgent.news from SiliconANGLE's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at siliconangle.com →

More in AI

Run your AI subscription 24 hours a day — use the quota you already pay for

Let me start with a question. Why did I fear development done by artificial intelligence? The answer is plain. AI can build software, and on top of that, it never rests.

  • AI never rests, operating 24/7 without labor laws or human limitations
  • First-mover advantage crucial for AI development, setting standards and pace
  • Unused subscription quota at night presents cost-efficient opportunity for small companies

More from Thursday 3 September →