Overview
The flagship creative text-to-speech model in Gemini 3.8, emphasizing studio-grade fidelity, expressive acting, regional accents, and long-form multi-turn stability.
The flagship creative text-to-speech model in Gemini 3.8, emphasizing studio-grade fidelity, expressive acting, regional accents, and long-form multi-turn stability.
Google announced Gemini 3.8 Flash TTS GA in the Gemini API on 2026-09-22; model ID `gemini-3.8-flash-tts`, targeting high-fidelity creative TTS with Voices, Voice Design, and Voice Replication.
Only published specifications are shown.
The flagship creative text-to-speech model in Gemini 3.8, emphasizing studio-grade fidelity, expressive acting, regional accents, and long-form multi-turn stability.
It extends Gemini 3.8 audio from live conversation into controllable, high-quality speech generation.
Gemini 3.8 Flash TTS reached GA in the Gemini API for high-fidelity creative text-to-speech, alongside Voices, Voice Design, and Voice Replication.
View sourceSelect a card to view the related model.
Copy this template, add the field, proposed value, and primary source, and share it with the archive maintainer.