Google 於 2026 年 9 月 23 日發表兩款文字轉語音模型:Flash TTS 側重角色聲音與細緻配音指引,Flash Lite TTS 則面向較低成本的大量音訊製作。[4][27]
Google 表示可選用超過 2,000 種現成聲音,並逐句指定說話方式;複製既有聲音須經同意驗證,與從文字描述創造新聲線不同。[29][23]
What are Google’s Gemini 3.8 Flash TTS and Flash-Lite TTS models, how do they differ in intended use, and what did Google announce about theAI-generated editorial illustration; not an image of the models’ interface.
AI 提示詞
Create a landscape editorial hero image for this Studio Global article: What are Google’s Gemini 3.8 Flash TTS and Flash-Lite TTS models, how do they differ in intended use, and what did Google announce about the. Article summary: Google announced Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS on September 23, 2026, as text-to-speech models for generating expressive spoken audio. Flash TTS is aimed at detailed creative direction and character . Topic tags: general, documentation, general web, user generated. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks,
Google 表示,語音複製設有同意驗證。上市報導也指出,生成音訊帶有 SynthID 數位浮水印,複製的聲音另附 C2PA 內容憑證,用於標示內容來源。這些是降低濫用風險的措施,不等於所有冒用情況都能杜絕。2334
在發表時公布的評測結果中,Flash TTS 於 Hume AI 的 Voice Design Benchmark 取得 71.4 分、排名第一;另有報導列出其口音模擬項目為 60.8 分。Google 也表示,Flash TTS 與 Flash-Lite TTS 包辦 Hume AI Overall Quality Index 前兩名。這些是發表時報告的基準測試成績,不能直接視為特定產品情境下的表現保證。1921
哪裡可以使用?
Google 表示,兩款模型自 9 月 23 日起在供開發者試用的 Google AI Studio 及用於程式串接的 Gemini API 推出。Gemini API 更新紀錄將新 TTS 模型及 Voices 端點列為正式提供。面向一般使用者的產品安排則是:Flash TTS 用於 Gemini Notebook,Flash-Lite TTS 用於 Google Vids;透過 Gemini Enterprise 使用兩款模型則標示為「即將推出」。這不表示所有聲音功能在推出當天已於每項產品開放。2932