Urgent.News

What's breaking now, across thousands of outlets.

AI

The AI may know when it is guessing

Imagine this: you ask an AI a question. It answers in a calm, confident voice. You trust it — until you discover that one detail was invented. The dangerous part of an AI hallucination is not only that it is wrong. It is that the answer often sounds completely sure. A new research paper asks a useful question: what if we could see the AI getting uncertain before it finished the sentence? The…

An intriguing research paper proposes a method called InnerExpert to detect when an AI model is uncertain about its responses. The technique works by examining an AI model's internal experts and identifying when they disagree on a particular piece of information. By spotting this disagreement early in the response creation process, InnerExpert aims to provide a warning signal, helping to prevent the model from generating potentially incorrect or misleading information.

The researchers tested InnerExpert on five datasets and two different Mixture-of-Experts (MoE) model designs. The results showed that the detector was effective at identifying risky answers, achieving an area under the receiver operating characteristic curve (AUROC) score of 0.91 for overall answers and 0.76 for individual words or tokens within answers. This score is higher than random guessing (0.50) but still falls short of perfect accuracy (1.00).

What sets InnerExpert apart from other methods of AI validation—such as sending the same question to multiple models or comparing AI-generated answers with external databases—is that it utilizes the model's own internal processes to identify uncertainty. This approach is cheaper and more efficient than the alternatives, as it doesn't require additional AI models or extensive comparisons.

While InnerExpert doesn't eliminate the risk of AI hallucinations entirely, it offers a proactive approach to flag potential errors before they are produced. This could be particularly useful in customer support, research assistance, or any application where the accuracy of AI-generated information is critical. However, it's important to note that InnerExpert does not guarantee the correctness of AI outputs and should be considered a tool to aid human review, not a substitute for thorough fact-checking.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in AI

Tech union to hold AI 'counter summit'

The Communications Workers' Union, which represents tech workers, is to hold an AI 'Counter Summit' today to highlight the impact of artificial intelligence on jobs.

More from Friday 9 October →