跳到正文
原文
GoogleAI· @GoogleAI · X·· 13 天前AI 评分64

Google 发布 Gemini 3.8 Flash TTS 和 Gemini 3.8 Flash-Lite TTS 语音模型

AI 导读

Google AI 发布 Gemini 3.8 Flash TTS 和 Gemini 3.8 Flash-Lite TTS 两款语音模型,支持 100+ 语言自定义声音或选用 2,000+ 现成声音,可指导多轮对话、逐行控制表达并加入 <laughs>、|mhm| 等自然提示,生成数小时一致无故障的音频。

正文 · 原文

We’re launching Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS ⚡️

Our most expressive audio models yet let you create custom voices across 100+ languages or pick from 2,000+ ready-to-use ones. You can direct back-and-forth conversations, guide the delivery line-by-line, and add natural cues like <laughs> or an active listening interjection like |mhm| all while generating hours of consistent, glitch-free audio.

Sounds pretty cool, right? So… how should you use them?

— Gemini 3.8 Flash TTS: Need to design bespoke vocal personas from scratch and with line-by-line level control? This is the model! Built for high-fidelity creative production like gaming, immersive audiobooks, and podcasts.

— Gemini 3.8 Flash-Lite TTS: Want the AI to automatically adjust its tone and pacing on the fly for near real-time voice agents? This is your engine! Built for cost-efficient scale, high-volume dubbing, and bulk audio creation.

来源:GoogleAI · x.com