Urgent.News

What's breaking now, across thousands of outlets.

More in Tech

Fourteen Speaker Encoders Heard the Same Voice. Their Error Rates Differed Five-Fold.

A speaker encoder turns a few seconds of speech into a vector, and the distance between two vectors is treated as an answer to "are these the same person." That number gets used for authentication…

  • Fourteen encoder models evaluated identical voice samples with varying error rates.
  • Error rates ranged from 0.047 to 0.233, indicating a five-fold discrepancy.
  • Emotional speech exacerbates error differences, with expressive speech showing higher error rates.

More from Monday 31 August →