Skip to main content
idapt
HomeCodeAI ModelsPricing
Sign inStart free
Fish Audio logo

Fish S1

NEW

by Fish Audio

Fish Audio's original S-series speech model.

Fish S1 pairs solid multilingual synthesis with Fish Audio's voice-cloning heritage, accepting any Fish marketplace voice id. Serves through OpenRouter's speech API.

More details

Fish S1 by Fish Audio is a standard-speed AI text-to-speech model. Fish S1 pairs solid multilingual synthesis with Fish Audio's voice-cloning heritage, accepting any Fish marketplace voice id. Serves through OpenRouter's speech API. Key capabilities include 274 voices across 12 languages, up to 4,096 characters per request.

Fish S1 is priced at $0.02 per 1,000 characters of input text. You can generate speech with Fish S1 on idapt.app alongside 200+ AI models in one workspace.

Voices

1 · 1 languages

Any of these voices can be paired with Fish S1. Pick a model for quality and speed, a voice for character and language.

Fish Default

Multilingual · Neutral

Pricing

Cost per 1,000 characters

$0.02

Typical paragraph (~500 chars)

$0.009

Billed on input text length

Capabilities

Speed

Standard: Balanced latency & quality

Emotion Control

Natural delivery only

Max Input

4,096 characters per request

Voices

1 voices · 1 languages

Generate speech with Fish S1View Details

Compare Fish S1 with

OpenAI logo
GPT-4o Mini TTSOpenAI
xAI logo
Grok Voice TTSxAI
Google AI Studio logo
Gemini 3.1 Flash TTS PreviewGoogle AI Studio
MiniMax logo
Speech 2.8 HDMiniMax
MiniMax logo
Speech 2.8 TurboMiniMax
MiniMax logo
Speech 02 HDMiniMax

Frequently Asked Questions

Who created Fish S1?
Fish S1 was created by Fish Audio.
How much does Fish S1 cost?
Fish S1 costs $0.02 per 1,000 characters of input text. A short paragraph of roughly 500 characters costs about $0.009.
How fast is Fish S1?
Fish S1 is classified as a "standard" speed model. It balances synthesis latency and audio quality for general-purpose use.
What voices and languages does Fish S1 support?
Fish S1 supports 274 curated voices spanning 12 languages, including English, Chinese, Korean, Japanese, Spanish, and Portuguese. Every voice in the catalog can be used with this model.
Does Fish S1 support emotion control?
Fish S1 does not expose explicit emotion control. Delivery follows the natural intonation of the selected voice.
How much text can Fish S1 synthesize at once?
Fish S1 accepts up to 4,096 characters of input text per request. Longer documents are split automatically across multiple requests.
How does Fish S1 compare to other TTS models?
Fish S1 is one of several text-to-speech models available on idapt. Compare it with GPT-4o Mini TTS, Grok Voice TTS, Gemini 3.1 Flash TTS Preview and more on the audio models directory.
Where can I use Fish S1?
You can generate speech with Fish S1 on idapt.app — in chat via the media tools, in hands-free voice mode, or through the public API. No separate API key required: idapt provides Fish S1 alongside 200+ AI models in one professional workspace.
Generate speech with Fish S1
  • Home
  • Pricing
  • AI Models
  • Image models
  • Voice models
  • Video models
  • Rankings
  • New models
  • Model status
  • Multi-Model Chat
  • Agents
  • Computers
  • Drive
  • Automations
  • AI Gateway
  • All features →
  • LLM cost calculator
  • Token counter
  • All free tools →
  • Blog
  • Use cases
  • Comparisons
  • Best of
  • Benchmarks
  • Demos
  • Changelog
  • Help center
  • FAQ
  • Privacy
  • Compare all models
  • Support
  • idapt Code
  • Developers
  • Quickstarts
  • API reference
  • API pricing
  • CLI
  • MCP
  • Downloads
  • Desktop
  • Badges and embeds
© idapt[email protected]TermsPrivacy PolicyLegal noticeReport content
X (Twitter)