Urgent.News

What's breaking now, across thousands of outlets.

AI

My speech-flaw detector flagged 41 false alarms a minute. One line of math fixed it.

I built Podium , a tool that compares your reading of a speech with a great delivery of the same text (JFK, Reagan) and tells you exactly where you drift and why: "14.3 syllables/s here vs 7.7 in the reference (+85%)", pinned to a time range. The first version worked beautifully on my test set. Then I gave it a different speaker, and it flagged 41 "flaws" per minute on a perfectly clean…

Podium is a tool that compares a person's reading of a speech with a great delivery of the same text and highlights where the reading deviates and why. Initially, the tool performed well on the test set. However, when applied to a different speaker, it generated 41 false alarms per minute on a perfect recording. The issue was addressed with a one-line fix and improved evaluation set-up.

The process involved aligning both recordings to the same word grid, building a dataset with exact answers by injecting flaws, and separating style from flaws. By determining what shouldn't count before deciding what should, the false alarms decreased significantly. The evaluation results showed that the tool detected flaws with a mean Intersection over Union (IoU) of 0.81, an overall F1 score of 0.51, and no false alarms on clean audio of the reference speaker.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in AI

Seu agente pode fazer deploy. Mas deveria?

SDD, Harness, Agents e Segurança em ambientes Multicloud São 16h52 de uma sexta-feira. Uma vulnerabilidade crítica acabou de ser identificada em uma aplicação em produção. O time precisa agir rápido.

  • Critical vulnerability found in production application running across multiple clouds
  • Developer seeks help for Agent to fix vulnerability, run tests, and prepare deployment

Outside Window: A Local AI Agent That Gets You Off the Screen

This is a submission for the Hacktoberfest Open-Source AI Challenge Week 1: Touch Grass . What I Built Outside Window is a privacy-first personal agent that helps someone take one realistic outdoor…

  • Outside Window is an open-source AI project
  • Encourages users to take breaks outside screens
  • Uses Gemma model for local inference

Concurrency-Aware Procurement: How Agentic Buyers Balance Parallel Negotiations Against Cancellation Risk

Agentic buyers can fork a procurement task into dozens of parallel negotiation threads. Spinning up another thread is cheap.

  • Agentic buyers can parallelize procurement tasks, but concurrency incurs costs.
  • CANO optimizer balances parallelism against cancellation risk through Monte Carlo simulations.
  • Key insights: diminishing returns on negotiators, convex quantile curve for price caps.

More from Tuesday 6 October →