NeFut Logo NeFut
中 Admin Login

[DeepMind] Overview of Gemini 3.8 Text-to-Speech Models

Published at: 2026-09-26 22:00 Last updated: 2026-09-28 00:49
#AI #Machine Learning #Neural

Gemini 3.8 Flash TTS and Flash‑Lite TTS are the newest voice‑generation models in the Gemini family. They support more than 100 languages and dialects and ship with a library of over 2,000 production‑ready voices. Flash TTS is built for deep creative direction, allowing developers to craft brand‑new characters from natural‑language prompts and control every line of speech. Flash‑Lite TTS is optimized for high‑volume, cost‑effective scaling, ideal for dubbing, podcasts, and expressive voice agents.

Key capabilities

Performance: On Hume AI’s Voice Design Benchmark, Flash TTS scores 71.4 overall (rank 1) and 60.8 on accent modeling (rank 1). Flash‑Lite TTS follows closely. In blind human preference tests on Voice Arena, both models rank top across major languages such as Japanese, Brazilian Portuguese, Vietnamese, Modern Standard Arabic, Mexican Spanish, and Hindi.

How to use

Partners like Figma, HeyGen, Linguana, Wondercraft, 99.co, and Ollang are already integrating the models for global dubbing, nuanced regional localization, and scalable conversational agents.

Review

Original Source: https://deepmind.google/blog/say-hello-to-gemini-38-text-to-speech/

Next: None
[h] Back to Home