Reference
Remy Reference/models/Gemini 3.1 Flash TTS
G

Gemini 3.1 Flash TTS

Gemini 3.1 Flash TTS, a text-to-speech model from Google.
speechModelOverride: { model: "gemini-3.1-flash-tts" }
Google·Released April 2026
Pricing
Input
$1.00 / 1M
Output
$20.00 / 1M
Specifications
Context window
16,384 tokens
Max output
16,384 tokens
Released
April 2026
Model string
gemini-3.1-flash-tts-preview
Parameters
Model options, passed in speechModelOverride.config.
ParameterTypeDefaultRange / options
voice
Prebuilt voice preset to use.
SelectKoreZephyrPuckCharonKoreFenrirLeda+24
styleInstruction
Optional natural-language direction for delivery (e.g. "Say cheerfully:", "Whisper softly:", "Narrate dramatically:"). Prepended to the input before synthesis. Leave blank for a neutral read. You can also embed expressive audio tags directly in your input text like [happy], [whisper], [laughing].
Prompt
Voices · 30
Zephyr
bright
Puck
upbeat
Charon
informative
Kore
firm
Fenrir
excitable
Leda
youthful
Orus
firm
Aoede
breezy
Callirrhoe
easy-going
Autonoe
bright
Enceladus
breathy
Iapetus
clear
Umbriel
easy-going
Algieba
smooth
Despina
smooth
Erinome
clear
Algenib
gravelly
Rasalgethi
informative
Laomedeia
upbeat
Achernar
soft
Alnilam
firm
Schedar
even
Gacrux
mature
Pulcherrima
forward
Achird
friendly
Zubenelgenubi
casual
Vindemiatrix
gentle
Sadachbia
lively
Sadaltager
knowledgeable
Sulafat
warm
Use this model
await mindstudio.textToSpeech({
  text: "...",
  speechModelOverride: {
    model: "gemini-3.1-flash-tts",
    config: {
      voice: "Zephyr",
    },
  },
});