Model index
Google DeepMind·Gemini

Gemini 3.8 Flash TTS

The flagship creative text-to-speech model in Gemini 3.8, emphasizing studio-grade fidelity, expressive acting, regional accents, and long-form multi-turn stability.

CurrentLanguageMultimodalWeights not released

Google announced Gemini 3.8 Flash TTS GA in the Gemini API on 2026-09-22; model ID `gemini-3.8-flash-tts`, targeting high-fidelity creative TTS with Voices, Voice Design, and Voice Replication.

MODEL NUMBER086
First public date
2026-09-22
Organization
Google DeepMind
Family
Gemini
01

Specifications

Only published specifications are shown.

Context
8,192 tokens↗
Maximum output
16,384 tokens↗
Input modalities
Text↗
Output modalities
Audio↗
Open weights
No
Status
Current↗
02

Overview

Overview

The flagship creative text-to-speech model in Gemini 3.8, emphasizing studio-grade fidelity, expressive acting, regional accents, and long-form multi-turn stability.

Why it mattered

It extends Gemini 3.8 audio from live conversation into controllable, high-quality speech generation.

03

Architecture & capabilities

Model capabilities

Audio generationVoice cloningMultilingual
04

Release history

General release

Gemini 3.8 Flash TTS reached GA in the Gemini API for high-fidelity creative text-to-speech, alongside Voices, Voice Design, and Voice Replication.

View source
05

Related models

Select a card to view the related model.

06

Sources

Suggest a correction

Copy this template, add the field, proposed value, and primary source, and share it with the archive maintainer.