{
  "id": 9457287,
  "title": "Gemini 3.8 text-to-speech",
  "url": "https://urgent.news/2026/09/23/gemini-3-8-text-to-speech",
  "topic": "ai",
  "section": "AI",
  "published": "2026-09-23T15:29:23.000Z",
  "source": {
    "name": "Hacker News Best",
    "slug": "hacker-news-best",
    "url": "https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-text-to-speech/"
  },
  "original_language": "en",
  "account": "Google has introduced two new text-to-speech models, dubbed Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, which bring a heightened level of customization and expressiveness to audio generation. These models empower creators, developers, and enterprises to craft unique, natural-sounding voices tailored to their specific needs, from character voices for audiobooks to brand ambassadors for podcasts.\n\nBoth TTS models offer robust control over voice delivery, securing top spots on leading quality benchmarks. Gemini 3.8 Flash TTS, for instance, outshines competitors with a Hume AI Voice Design Benchmark score of 71.4, while Gemini 3.8 Flash-Lite TTS ranks second in overall quality.\n\nThese advancements are complemented by the Gemini Audio family, which already includes models like 3.5 Live Translate, 3.5 Transcribe, 3.8 Live, and 3.8 Live Extended Thinking. With the ability to scale from 30 original voices to an infinite library, users can now create richly expressive audio experiences, whether they're building an entirely new character or fine-tuning a consistent brand voice.\n\nTo ensure responsible use, Google has implemented strict safeguards. For voice replication, users must provide a verbal consent recording from the voice owner, and every audio clip generated by Gemini Audio models is watermarked with SynthID to prevent misinformation.\n\nThese new models are now available in Google AI Studio, allowing developers to build and deploy speech generation experiences with ease. Companies like Figma, HeyGen, Linguana, Wondercraft, 99.co, and Ollang have already integrated these TTS models, enhancing global dubbing, localization, and conversational voice agents. However, it's worth noting that voice replication through AI Studio is currently unavailable in certain regions such as Illinois, Texas, EEA, UK, Switzerland, and India.",
  "summary": "Article URL: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-text-to-speech/ Comments URL: https://news.ycombinator.com/item?id=49817615 Points: 249 # Comments: 121",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}