MODEL / EXPRESSIVE SPEECH

Gemini 3.8 Flash TTS

Give a written script a voice, a pace, and a point of view. Gemini 3.8 Flash TTS turns narration or a two-person exchange into speech you can listen to and download. Start with a short passage below, choose a voice, and review the credit estimate before generating.

Model

Script

0 characters / 5,000

Single voice

Voice sample uses neutral settings.

Advanced controls
Estimated credits1.2

Flash · English welcome

Generated for this site. Script excerpt: “Hello! Welcome to Gemini TTS.”

Direct the delivery with Gemini 3.8 Flash TTS

A sentence can welcome someone, explain a difficult idea, or leave a question hanging in the air. Start by deciding which of those jobs your recording needs to do. The voice gives you a starting character; the script and direction tell it how to approach the moment. Use a short voice description such as “warm and measured, with a curious finish” instead of stacking a long list of competing moods.

Gemini 3.8 Flash TTS supports the controls exposed in this studio: voice, accent, style, pace, a voice description, and scene direction. You can leave optional controls at their defaults. A plain sentence is often the most useful first test because it lets you hear the selected voice without several instructions pulling in different directions. Change one setting, generate another short take, and compare the same passage.

Build a script that works when spoken

Read your text aloud before generating. Break long sentences where a person would naturally breathe, write unfamiliar abbreviations in a clearer form, and replace written shortcuts that sound awkward when spoken. Punctuation guides phrasing, while the supported inline tags provide specific cues for a pause, breath, laugh, or sigh. Put a cue where it belongs in the sentence, rather than collecting instructions at the end.

A short paragraph is enough to judge whether your opening sounds right. Listen for the pronunciation of names, the stress on important words, and the transition between sentences. If one line needs work, revise that line first. The generator accepts up to 5,000 characters per request; a shorter test leaves room to experiment before committing to a longer passage. Tags are part of the submitted script and contribute to its length.

A quiet story opening

A measured narrator introduces a scene, then lets a pause carry the change in mood.

The last train had already left when Mara reached the station. <short pause> On the empty bench sat a small blue envelope. It had her name on it. She looked toward the clock, then back at the envelope. Somewhere behind the ticket office, a telephone began to ring.

Voice direction: Calm storyteller. Warm, measured delivery; let the final sentence become quietly curious.

Pauses and vocal cues

Try a small number of supported tags where the spoken moment calls for them.

I thought I had forgotten the keys. <sigh> Then I checked the pocket of my coat. <short pause> There they were, right beside the note reminding me to bring them. <laugh> At least the note worked.

Voice direction: Light, conversational storytelling. Keep the reactions understated.

Two voices with Gemini 3.8 Flash TTS

For a conversation, switch to two speakers and choose a distinct voice for each participant. Assign every turn to speaker one or speaker two. A curious host and a measured guest are a useful starting pair: the contrast makes it easier to follow the conversation without exaggerating either performance. Scene direction can describe their shared situation, while each voice description describes only that speaker.

Keep an exchange focused on one topic. Start with a question, give the second speaker space to answer, and use a follow-up that responds to the answer. The studio submits the ordered turns as one dialogue and returns one audio result. It does not provide separate speaker tracks or an automatic video edit. For a reusable starting point, open the interview template and replace its lines with your own exchange.

A short interview

A host asks a focused question and a guest answers with a practical example.

Speaker 1: What helped you finish your first project? Speaker 2: Making the first version smaller. I chose one problem and tested it with a friend. Speaker 1: What changed after that conversation? Speaker 2: I rewrote the opening. Once the purpose was clear, the rest was much easier to explain.

Scene direction: A relaxed interview. The host is curious; the guest answers thoughtfully without rushing.

Listen to a real model sample

The recording on this page is an English welcome generated for this site with Flash. It demonstrates one short take, not every voice or every possible delivery. The voice library contains neutral previews of all thirty available voices. Use those previews to narrow your choice, then test your own text in the studio: a voice that suits a friendly greeting may need different direction for a technical explanation.

A template is editable text and direction, not a pre-generated recording. Applying one does not use generation credits. You can change the voice, remove a tag, or rewrite a line before submitting. Playback and WAV download become available when your task completes, and completed recordings can also be found in your library.

Choosing between Flash and Flash-Lite

Use Gemini 3.8 Flash TTS when you want to explore a performance with narration, character dialogue, or changes in tone. Flash-Lite is also available in the same interface and has a lower configured credit rate. Both models use the current voice and script controls. The best comparison is a short passage generated with the same voice and instructions, followed by a careful listen to the result.

The Flash-Lite page contains the detailed comparison and current credit examples. There is no need to rewrite a script to switch models, but switching does not regenerate an existing recording. Each submitted take is a new generation request. Check the estimate after changing the model or text, and keep your preferred WAV file before continuing with a new variation.

Questions before you start

Can I try Gemini 3.8 Flash TTS without an API key?

Yes. Use this browser application with your account and available credits; you do not enter an API key in the studio. You can browse templates and listen to voice previews before signing in. Generating a new recording requires authentication and enough credits for the selected model and script.

Does Flash support two speakers here?

Yes. Select Two speakers, assign a voice to each role, and write the conversation as ordered turns. The combined submitted text must fit the character limit. Use the dedicated dialogue page for interview, lesson, and story examples.

What can I download?

Completed tasks offer a WAV download. Listen to the entire result before using it in an edit. This page does not promise MP3 conversion, exact timestamps, a fixed audio duration, or separate tracks for individual speakers.

Is this the official Google service?

GeminiTTS is an independent application using Gemini speech models. Its account, credit balance, and prices belong to this application. Check the pricing page for these service rates and the official model documentation for Google’s model information.