Google 发布 Gemini 3.8 Flash TTS,可用文字描述自定义语音
Google's new Flash TTS models let you design AI voices from scratch using text descriptions
Google 新出的 Gemini 3.8 Flash TTS 能用一句话描述就造出音色,还能一段剧本生成双人对话,做播客和配音的可以试试。
Google 推出 Gemini 3.8 Flash TTS 和 Flash-Lite TTS 两款语音合成模型,支持超过 100 种语言。Flash TTS 可以根据文字描述生成全新音色。两个模型都支持对单句添加舞台指示,并从一份剧本生成双人对话。语音克隆功能只需 30 秒采样即可建立声音档案。
Google's new Flash TTS models let you design AI voices from scratch using text descriptions
Google is introducing two new text-to-speech models, Gemini 3.8 Flash TTS and Flash-Lite TTS, which support more than 100 languages. Flash TTS can create new voices from text descriptions, and both models let users add stage directions to individual lines and generate two-voice dialogue from a single script. A voice cloning feature can build a voice profile from a 30-second sample. The article Google's new Flash TTS models let you design AI voices from scratch using text descriptions appeared first on The Decoder .