Skip to main content
idapt
HomeCodeAI ModelsPricing
Sign inStart free trial
Google AI Studio logo

Gemini 3.1 Flash TTS Preview

by Google AI Studio

Google's newest controllable Gemini text-to-speech preview model.

Gemini 3.1 Flash TTS Preview is tuned for low-latency, steerable speech generation with single-speaker and two-speaker output, making it a strong choice for narrated workflows and dialog-style audio.

More details

Gemini 3.1 Flash TTS Preview by Google AI Studio is a fast-speed AI text-to-speech model. Gemini 3.1 Flash TTS Preview is tuned for low-latency, steerable speech generation with single-speaker and two-speaker output, making it a strong choice for narrated workflows and dialog-style audio. Key capabilities include 78 voices across 7 languages, emotion control, up to 15,000 characters per request.

Gemini 3.1 Flash TTS Preview is priced at $0.04 per 1,000 characters of input text. You can generate speech with Gemini 3.1 Flash TTS Preview on idapt.app alongside 200+ AI models in one workspace.

Voices

30 · 1 languages

Any of these voices can be paired with Gemini 3.1 Flash TTS Preview. Pick a model for quality and speed, a voice for character and language.

Zephyr

Multilingual · Neutral

Puck

Multilingual · Neutral

Charon

Multilingual · Neutral

Kore

Multilingual · Neutral

Fenrir

Multilingual · Neutral

Leda

Multilingual · Neutral

Orus

Multilingual · Neutral

Aoede

Multilingual · Neutral

Callirrhoe

Multilingual · Neutral

Autonoe

Multilingual · Neutral

Enceladus

Multilingual · Neutral

Iapetus

Multilingual · Neutral

Umbriel

Multilingual · Neutral

Algieba

Multilingual · Neutral

Despina

Multilingual · Neutral

Erinome

Multilingual · Neutral

Algenib

Multilingual · Neutral

Rasalgethi

Multilingual · Neutral

Laomedeia

Multilingual · Neutral

Achernar

Multilingual · Neutral

Alnilam

Multilingual · Neutral

Schedar

Multilingual · Neutral

Gacrux

Multilingual · Neutral

Pulcherrima

Multilingual · Neutral

Achird

Multilingual · Neutral

Zubenelgenubi

Multilingual · Neutral

Vindemiatrix

Multilingual · Neutral

Sadachbia

Multilingual · Neutral

Sadaltager

Multilingual · Neutral

Sulafat

Multilingual · Neutral

Pricing

Cost per 1,000 characters

$0.04

Typical paragraph (~500 chars)

$0.02

Billed on input text length

Capabilities

Speed

Fast: Low-latency, real-time ready

Emotion Control

Steer happy / sad / angry / neutral delivery

Max Input

15,000 characters per request

Voices

30 voices · 1 languages

Generate speech with Gemini 3.1 Flash TTS PreviewView Details

Compare Gemini 3.1 Flash TTS Preview with

OpenAI logo
GPT-4o Mini TTSOpenAI
xAI logo
Grok Voice TTSxAI
MiniMax logo
Speech 2.8 HDMiniMax
MiniMax logo
Speech 2.8 TurboMiniMax
MiniMax logo
Speech 02 HDMiniMax

Frequently Asked Questions

Who created Gemini 3.1 Flash TTS Preview?
Gemini 3.1 Flash TTS Preview was created by Google AI Studio.
How much does Gemini 3.1 Flash TTS Preview cost?
Gemini 3.1 Flash TTS Preview costs $0.04 per 1,000 characters of input text. A short paragraph of roughly 500 characters costs about $0.02.
How fast is Gemini 3.1 Flash TTS Preview?
Gemini 3.1 Flash TTS Preview is classified as a "fast" speed model. It is tuned for low-latency synthesis, making it well suited to interactive and real-time use.
What voices and languages does Gemini 3.1 Flash TTS Preview support?
Gemini 3.1 Flash TTS Preview supports 78 curated voices spanning 7 languages, including English, Chinese, Korean, Japanese, Spanish, and Portuguese. Every voice in the catalog can be used with this model.
Does Gemini 3.1 Flash TTS Preview support emotion control?
Yes, Gemini 3.1 Flash TTS Preview supports emotion control. You can steer the delivery toward tones such as happy, sad, angry, or neutral for more expressive speech.
How much text can Gemini 3.1 Flash TTS Preview synthesize at once?
Gemini 3.1 Flash TTS Preview accepts up to 15,000 characters of input text per request. Longer documents are split automatically across multiple requests.
How does Gemini 3.1 Flash TTS Preview compare to other TTS models?
Gemini 3.1 Flash TTS Preview is one of several text-to-speech models available on idapt. Compare it with GPT-4o Mini TTS, Grok Voice TTS, Speech 2.8 HD and more on the audio models directory.
Where can I use Gemini 3.1 Flash TTS Preview?
You can generate speech with Gemini 3.1 Flash TTS Preview on idapt.app — in chat via the media tools, in hands-free voice mode, or through the public API. No separate API key required: idapt provides Gemini 3.1 Flash TTS Preview alongside 200+ AI models in one professional workspace.
Generate speech with Gemini 3.1 Flash TTS Preview
  • Home
  • Pricing
  • AI Models
  • Image models
  • Voice models
  • Video models
  • Rankings
  • New models
  • Model status
  • Multi-Model Chat
  • Agents
  • Computers
  • Drive
  • Automations
  • AI Gateway
  • All features →
  • LLM cost calculator
  • Token counter
  • All free tools →
  • Blog
  • Use cases
  • Comparisons
  • Best of
  • Benchmarks
  • Changelog
  • Help center
  • FAQ
  • Privacy
  • Compare all models
  • Support
  • idapt Code
  • Developers
  • Quickstarts
  • API reference
  • API pricing
  • CLI
  • MCP
  • Downloads
  • Desktop
  • Badges and embeds
© idapt[email protected]TermsPrivacy PolicyLegal noticeReport content
X (Twitter)