Google 发布 Gemini 3.8 Flash TTS 与 Flash-Lite TTS:声音、复刻与上线范围
Google 于 2026 年 9 月 23 日发布两款文本转语音模型:Flash TTS 偏重角色声音和细致的表演指导,Flash Lite TTS 面向注重成本的大批量配音及语音助手场景。[4][27] Google 称提供超过 2000 种现成声音,并支持逐句指导语气、口音和语速;复刻已有说话者的声音则需要验证其同意。[29][28][23] 两款模型开始在 Google AI Studio 和 Gemini API 推出;Gemini Enterprise 的接入被列为“即将推出”。[29][32]
Google 于 2026 年 9 月 23 日发布两款文本转语音模型:Flash TTS 偏重角色声音和细致的表演指导,Flash Lite TTS 面向注重成本的大批量配音及语音助手场景。[4][27]
Google 称提供超过 2000 种现成声音,并支持逐句指导语气、口音和语速;复刻已有说话者的声音则需要验证其同意。[29][28][23]
两款模型开始在 Google AI Studio 和 Gemini API 推出;Gemini Enterprise 的接入被列为“即将推出”。[29][32]
What are Google’s Gemini 3.8 Flash TTS and Flash-Lite TTS models, how do they differ in intended use, and what did Google announce about theAI-generated editorial illustration; not an image of the models’ interface.
AI 提示
Create a landscape editorial hero image for this Studio Global article: What are Google’s Gemini 3.8 Flash TTS and Flash-Lite TTS models, how do they differ in intended use, and what did Google announce about the. Article summary: Google announced Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS on September 23, 2026, as text-to-speech models for generating expressive spoken audio. Flash TTS is aimed at detailed creative direction and character . Topic tags: general, documentation, general web, user generated. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks,
Google 表示,声音复刻设有同意验证。发布报道还提到,生成音频带有 SynthID 水印,复刻声音附带 C2PA 内容凭证。这些是识别和约束用途的措施,不能理解为已消除所有滥用风险。2334
据发布时披露的测试结果,Flash TTS 在 Hume AI 的 Voice Design Benchmark 中取得 71.4 分、排名第一;另一份报道列出的口音建模成绩为 60.8 分。Google 还称,Flash TTS 和 Flash-Lite TTS 包揽 Hume AI Overall Quality Index 的前两名。这些是已公布的基准测试成绩,并不等于两款模型在具体项目中的效果已经得到独立验证。1921
目前可以在哪里使用?
Google 表示,两款模型自 9 月 23 日起在面向开发者的 Google AI Studio 和 Gemini API 推出。Gemini API 更新日志将新 TTS 模型及 Voices 接口列为正式发布。面向具体产品,Google 指定 Flash TTS 用于 Gemini Notebook、Flash-Lite TTS 用于 Google Vids;通过 Gemini Enterprise 接入则被列为“即将推出”。某款模型进入产品,不代表该产品在发布当天就具备其全部声音功能。2932