Urgent.News

What's breaking now, across thousands of outlets.

Tech

Open source tool distills Jev so you can run it locally

Jevstiller targets 98% agreement by learning familiar requests on your hardware while sending uncertain and audited queries upstream

Open source tool distills Jev so you can run it locally

Jevstiller is an open-source project that aims to create a local version of the Jev AI model, allowing it to run on a user's own hardware. This distillation process creates a smaller, faster model that can handle certain types of queries more efficiently than the original Jev AI. By answering questions locally, Jevstiller reduces the need for expensive Jev token usage and speeds up response times.

The local model achieves an impressive 98% agreement with Jev, while still delegating more complex queries to the main Jev servers. The Jevstiller team describes it as a "cache" that sits in front of Jev, determining when to route requests to the main servers based on the model's confidence in its own answers. An audit system continuously checks the local model's responses against Jev's, maintaining an agreement rate of 98%.

If the agreement rate drops below the target, the system automatically redirects more requests to Jev and retraining begins. Despite its high agreement rate, Jevstiller does not guarantee accuracy, as AI models can still make mistakes. The Jevstiller team's experiments demonstrated that the local model can quickly adjust to changes in Jev's behavior, maintaining its performance. Users can learn how to set up Jevstiller from the project's documentation.

Written by urgent.news from The Register's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Also reported by 1 other outlet

Read the original at theregister.com →

More in Tech

Meet Mind

Introduction Have you ever entered a meeting and struggled to remember what was discussed in the previous conversation? Important details such as requirements, concerns, preferences, and follow-ups…

An Agent Fleet, Sorted By Job

"We have agents" is about as informative as "we have software". The question is what they do, when they run and who notices when one stalls. Ours grew into a fleet step by step.

  • Agent fleets are organized groups of specialized software agents
  • Leadership agents summarize agent work and generate reports
  • Review agents provide unbiased, critical analysis of agent work

Backend Uptime Alerts from One-Minute Metrics API Polling and Failure Evidence

A one-minute poller is only useful if it preserves enough evidence to explain an outage after the alert fires. For a gaming backend, the practical choice is to evaluate a low-cardinality availability…

  • One-Minute Metrics API polls backend uptime at one-minute intervals.
  • Separate detection and diagnosis for incident records with timestamps and error context.
  • Node.js worker handles uptime polling with state machine and business policy thresholds.

I Built a Customer Support Agent That Remembers With Hindsight

The most frustrating thing about customer support is not always the problem itself. It is having to explain the same problem again after someone already knows the history.

  • Priya's Wi-Fi drops prompted Netgear router troubleshooting
  • RecallAI uses Hindsight for long-term memory in customer support
  • Customer isolation ensures Priya's router info doesn't mix with others

More from Tuesday 29 September →