← 返回事件
持续讨论AI

Google's new Flash TTS models let you design AI voices from scratch using text descriptions

图:The Decoder

发生了什么

Google is introducing two new text-to-speech models, Gemini 3.8 Flash TTS and Flash-Lite TTS, which support more than 100 languages. Flash TTS can create new voices from text descriptions, and both models let users add stage directions to individual lines and generate two-voice dialogue from a single script. A voice cloning feature can build a voice profile from a 30-second sample.

摘要按规则整理自下方来源原文

来源