Skip to main content
idapt
HomeCodeAI ModelsPricing
Sign inStart free
Qwen logo

Qwen3.0 Audio TTS Plus

NEW

by Qwen

Qwen's Plus-tier text-to-speech model with rich Chinese voices.

Qwen Audio 3.0 TTS Plus renders expressive Mandarin speech with natural prosody through its Longan voice catalog. Serves through OpenRouter's speech API.

More details

Qwen3.0 Audio TTS Plus by Qwen is a standard-speed AI text-to-speech model. Qwen Audio 3.0 TTS Plus renders expressive Mandarin speech with natural prosody through its Longan voice catalog. Serves through OpenRouter's speech API. Key capabilities include 274 voices across 12 languages, up to 4,096 characters per request.

Qwen3.0 Audio TTS Plus is priced at $0.02 per 1,000 characters of input text. You can generate speech with Qwen3.0 Audio TTS Plus on idapt.app alongside 200+ AI models in one workspace.

Voices

2 · 1 languages

Any of these voices can be paired with Qwen3.0 Audio TTS Plus. Pick a model for quality and speed, a voice for character and language.

Longan Lingxin

Chinese · Female

Longan Lufeng

Chinese · Male

Pricing

Cost per 1,000 characters

$0.02

Typical paragraph (~500 chars)

$0.01

Billed on input text length

Capabilities

Speed

Standard: Balanced latency & quality

Emotion Control

Natural delivery only

Max Input

4,096 characters per request

Voices

2 voices · 1 languages

Generate speech with Qwen3.0 Audio TTS PlusView Details

Compare Qwen3.0 Audio TTS Plus with

OpenAI logo
GPT-4o Mini TTSOpenAI
xAI logo
Grok Voice TTSxAI
Google AI Studio logo
Gemini 3.1 Flash TTS PreviewGoogle AI Studio
MiniMax logo
Speech 2.8 HDMiniMax
MiniMax logo
Speech 2.8 TurboMiniMax
MiniMax logo
Speech 02 HDMiniMax

Frequently Asked Questions

Who created Qwen3.0 Audio TTS Plus?
Qwen3.0 Audio TTS Plus was created by Qwen.
How much does Qwen3.0 Audio TTS Plus cost?
Qwen3.0 Audio TTS Plus costs $0.02 per 1,000 characters of input text. A short paragraph of roughly 500 characters costs about $0.01.
How fast is Qwen3.0 Audio TTS Plus?
Qwen3.0 Audio TTS Plus is classified as a "standard" speed model. It balances synthesis latency and audio quality for general-purpose use.
What voices and languages does Qwen3.0 Audio TTS Plus support?
Qwen3.0 Audio TTS Plus supports 274 curated voices spanning 12 languages, including English, Chinese, Korean, Japanese, Spanish, and Portuguese. Every voice in the catalog can be used with this model.
Does Qwen3.0 Audio TTS Plus support emotion control?
Qwen3.0 Audio TTS Plus does not expose explicit emotion control. Delivery follows the natural intonation of the selected voice.
How much text can Qwen3.0 Audio TTS Plus synthesize at once?
Qwen3.0 Audio TTS Plus accepts up to 4,096 characters of input text per request. Longer documents are split automatically across multiple requests.
How does Qwen3.0 Audio TTS Plus compare to other TTS models?
Qwen3.0 Audio TTS Plus is one of several text-to-speech models available on idapt. Compare it with GPT-4o Mini TTS, Grok Voice TTS, Gemini 3.1 Flash TTS Preview and more on the audio models directory.
Where can I use Qwen3.0 Audio TTS Plus?
You can generate speech with Qwen3.0 Audio TTS Plus on idapt.app — in chat via the media tools, in hands-free voice mode, or through the public API. No separate API key required: idapt provides Qwen3.0 Audio TTS Plus alongside 200+ AI models in one professional workspace.
Generate speech with Qwen3.0 Audio TTS Plus
  • Home
  • Pricing
  • AI Models
  • Image models
  • Voice models
  • Video models
  • Rankings
  • New models
  • Model status
  • Multi-Model Chat
  • Agents
  • Computers
  • Drive
  • Automations
  • AI Gateway
  • All features →
  • LLM cost calculator
  • Token counter
  • All free tools →
  • Blog
  • Use cases
  • Comparisons
  • Best of
  • Benchmarks
  • Demos
  • Changelog
  • Help center
  • FAQ
  • Privacy
  • Compare all models
  • Support
  • idapt Code
  • Developers
  • Quickstarts
  • API reference
  • API pricing
  • CLI
  • MCP
  • Downloads
  • Desktop
  • Badges and embeds
© idapt[email protected]TermsPrivacy PolicyLegal noticeReport content
X (Twitter)