{
  "id": 5880531,
  "title": "OmniVoice TTS model 600+ languages, voice cloning from 3-second clips, free open source.",
  "url": "https://urgent.news/2026/09/06/omnivoice-tts-600-3",
  "topic": "ai",
  "section": "AI",
  "published": "2026-09-06T01:18:54.000Z",
  "source": {
    "name": "Dev.to",
    "slug": "dev-to",
    "url": "https://dev.to/sarantoon/omnivoice-omedl-tts-600-phaasaa-okhlnesiiyngcchaakkhlip-3-winaathii-epidchrsfrii-2ane"
  },
  "original_language": "th",
  "account": "The k2-fsa team has introduced OmniVoice, a zero-shot text-to-speech model that supports over 600 languages. This model uses a new diffusion language model architecture and allows for voice cloning and design. OmniVoice can clone a voice from a 3-15 second audio clip and generate speech in a specific accent or style. The model is open-source and available on GitHub and Hugging Face.",
  "summary": "OmniVoice TTS model 600+ languages, clone voices from 3-second clips, free open source. The k2-fsa team launched OmniVoice, a zero-shot text-to-speech model that supports more than 600 languages, which is the widest language coverage among existing zero-shot TTS models. It uses a new diffusion language model architecture and supports both voice cloning and voice design.",
  "key_points": [
    "OmniVoice TTS model generates speech in over 600 languages",
    "Developed by k2-fsa with diffusion language model architecture",
    "Supports voice cloning and design from 3-second video clips"
  ],
  "editors_take": "This development lets developers generate high-quality speech in over 600 languages, including Thai, on local devices without cloud reliance, but raises ethical considerations around voice cloning.",
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}