Skip to main content
idapt
HomeAI ModelsExploreResourcesPricing
Sign inUse for Free
Google AI Studio logo

Gemini 3.8 Flash TTS

NEW

by Google AI Studio

Google's newest controllable Gemini text-to-speech model.

Gemini 3.8 Flash TTS is tuned for low-latency, steerable speech generation with single-speaker and two-speaker output, making it a strong choice for narrated workflows and dialog-style audio.

More details

Gemini 3.8 Flash TTS by Google AI Studio is a fast-speed AI text-to-speech model. Gemini 3.8 Flash TTS is tuned for low-latency, steerable speech generation with single-speaker and two-speaker output, making it a strong choice for narrated workflows and dialog-style audio. Key capabilities include 278 voices across 12 languages, emotion control, up to 15,000 characters per request.

Gemini 3.8 Flash TTS is priced at $0.04 per 1,000 characters of input text. You can generate speech with Gemini 3.8 Flash TTS on idapt.app alongside 200+ AI models in one workspace.

Voices

30 · 1 languages

Any of these voices can be paired with Gemini 3.8 Flash TTS. Pick a model for quality and speed, a voice for character and language.

Zephyr

Multilingual · Neutral

Puck

Multilingual · Neutral

Charon

Multilingual · Neutral

Kore

Multilingual · Neutral

Fenrir

Multilingual · Neutral

Leda

Multilingual · Neutral

Orus

Multilingual · Neutral

Aoede

Multilingual · Neutral

Callirrhoe

Multilingual · Neutral

Autonoe

Multilingual · Neutral

Enceladus

Multilingual · Neutral

Iapetus

Multilingual · Neutral

Umbriel

Multilingual · Neutral

Algieba

Multilingual · Neutral

Despina

Multilingual · Neutral

Erinome

Multilingual · Neutral

Algenib

Multilingual · Neutral

Rasalgethi

Multilingual · Neutral

Laomedeia

Multilingual · Neutral

Achernar

Multilingual · Neutral

Alnilam

Multilingual · Neutral

Schedar

Multilingual · Neutral

Gacrux

Multilingual · Neutral

Pulcherrima

Multilingual · Neutral

Achird

Multilingual · Neutral

Zubenelgenubi

Multilingual · Neutral

Vindemiatrix

Multilingual · Neutral

Sadachbia

Multilingual · Neutral

Sadaltager

Multilingual · Neutral

Sulafat

Multilingual · Neutral

Pricing

Cost per 1,000 characters

$0.04

Typical paragraph (~500 chars)

$0.02

Billed on input text length

Capabilities

Speed

Fast: Low-latency, real-time ready

Emotion Control

Steer happy / sad / angry / neutral delivery

Max Input

15,000 characters per request

Voices

30 voices · 1 languages

Generate speech with Gemini 3.8 Flash TTSView Details

Compare Gemini 3.8 Flash TTS with

ByteDance Seed logo
Seed Audio 1.0ByteDance Seed
OpenAI logo
GPT-4o Mini TTSOpenAI
xAI logo
Grok Voice TTSxAI
Google AI Studio logo
Gemini 3.1 Flash TTS PreviewGoogle AI Studio
MiniMax logo
Speech 2.8 HDMiniMax
MiniMax logo
Speech 2.8 TurboMiniMax

Frequently Asked Questions

Who created Gemini 3.8 Flash TTS?
Gemini 3.8 Flash TTS was created by Google AI Studio.
How much does Gemini 3.8 Flash TTS cost?
Gemini 3.8 Flash TTS costs $0.04 per 1,000 characters of input text. A short paragraph of roughly 500 characters costs about $0.02.
How fast is Gemini 3.8 Flash TTS?
Gemini 3.8 Flash TTS is classified as a "fast" speed model. It is tuned for low-latency synthesis, making it well suited to interactive and real-time use.
What voices and languages does Gemini 3.8 Flash TTS support?
Gemini 3.8 Flash TTS supports 278 curated voices spanning 12 languages, including English, Chinese, Korean, Japanese, Spanish, and Portuguese. Every voice in the catalog can be used with this model.
Does Gemini 3.8 Flash TTS support emotion control?
Yes, Gemini 3.8 Flash TTS supports emotion control. You can steer the delivery toward tones such as happy, sad, angry, or neutral for more expressive speech.
How much text can Gemini 3.8 Flash TTS synthesize at once?
Gemini 3.8 Flash TTS accepts up to 15,000 characters of input text per request. Longer documents are split automatically across multiple requests.
How does Gemini 3.8 Flash TTS compare to other TTS models?
Gemini 3.8 Flash TTS is one of several text-to-speech models available on idapt. Compare it with Seed Audio 1.0, GPT-4o Mini TTS, Grok Voice TTS and more on the audio models directory.
Where can I use Gemini 3.8 Flash TTS?
You can generate speech with Gemini 3.8 Flash TTS on idapt.app — in chat via the media tools, in hands-free voice mode, or through the public API. No separate API key required: idapt provides Gemini 3.8 Flash TTS alongside 200+ AI models in one professional workspace.
Generate speech with Gemini 3.8 Flash TTS
  • Home
  • Pricing
  • AI Models
  • Skills
  • Image models
  • Voice models
  • Video models
  • Rankings
  • Multi-Model Chat
  • Agents
  • Computers
  • Drive
  • Automations
  • AI Gateway
  • All features →
  • Blog
  • About
  • Comparisons
  • Benchmarks
  • Model Match
  • Demos
  • Help center
  • FAQ
  • Privacy
  • Support
  • Developers
  • Quickstarts
  • API reference
  • API pricing
  • CLI
  • MCP
  • Migrate
  • Downloads
  • Desktop
© idapt[email protected]TermsPrivacy PolicyLegal noticeReport content
X (Twitter)Discord