Urgent.News

What's breaking now, across thousands of outlets.

Tech

Zet: an open-source layer on Laya that knows when to ask a human

Confidence scores aren't error rates. Zet sits on top of Laya and marks every answer sure or unsure, using conformal prediction sets and a Learn-then-Test threshold calibrated per language, so the answers it automates stay within an error budget you choose (say 5%). On MASSIVE (human-labeled, 300 examples per language, 5% budget), it automated 84% of English and 80% of Swedish requests on a…

Zet is an open-source tool designed to work alongside Laya, a language model. Unlike traditional models, Zet does not provide confidence scores as error rates. Instead, it uses conformal prediction sets and a Learn-then-Test threshold that is calibrated per language. This allows Zet to ensure its automated answers remain within a user-defined error budget, such as 5%.

In extensive testing on a task with 300 examples per language and a 5% error budget, Zet automated 84% of English and 80% of Swedish requests on a 6-scenario task, with no errors in the 75 sure English answers. However, as the task complexity increased to 18 scenarios, Zet ceased automation, yet its prediction sets still contained the correct answer 97-99% of the time.

This capability to refuse when the task becomes too challenging is Zet's key feature. The tool runs locally on ONNX Runtime without requiring PyTorch, and offers a web interface for creating tasks, uploading labeled examples, calibrating, and reviewing answers that are uncertain. Zet is licensed under the Apache-2.0 license and is built on Laya, though it is not endorsed by the authors of Laya.

More information can be found on its website at https://yoosseph.github.io/Zet/ or its repository at https://github.com/Yoosseph/Zet.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in Tech

Six Signatures, Zero Laws: What the White House Super Intelligence Accord Actually Requires

If the entire safety framework for the most powerful technology on the planet were a document, how long would you expect it to be? As of Tuesday, the answer is two pages.

  • White House Accord signed by President Trump and six AI companies on September 29, 2026
  • Accord outlines voluntary commitments for safe AI development without legal enforcement
  • Main components include internal controls, watch team, external auditor, board oversight

What Is Decisioning Infrastructure for Consumer Platforms?

Consumer platforms make thousands or millions of decisions every second about what users see. A social platform decides which posts appear in a feed. A marketplace decides which listings appear first.

  • Decisioning infrastructure separates candidate generation from ranking and business logic.
  • It determines final item ordering, monetization placements, and records decisions.
  • Incorporates relevance, personalization, context, business rules, and diversity.

How to Find and Fix Memory Leaks in .NET Projects

Memory leaks in .NET applications can be confusing. After all, .NET has a Garbage Collector (GC) that automatically removes objects that are no longer needed.

  • Memory leaks occur due to lingering references in .NET applications despite garbage collection.
  • Employ dotnet-dump to capture memory dumps and analyze them for objects consuming excessive memory.

Freight Contract PDF Archive: Compress Copies, Store Originals for Signature Fidelity

Short answer: compress a PDF archive's access copies, but store signed originals unchanged. Storage cost matters; signature fidelity decides which bytes are authoritative.

  • Compress PDF contract copies while retaining original signed versions
  • Archive ownership determines proof of signed contracts
  • Align archive choice with server-side logistics template ownership

More from Wednesday 30 September →