Qwen-Audio-3.0-TTS Flash

Alibaba's low-latency Qwen-Audio-3.0 text-to-speech model optimized for real-time interaction across 16 languages. Generates speech via the /v1/audio/speech endpoint.

qwen-audio-3.0-tts-flash
STABLEGet StartedView uptime
20,000 context
Starting at $0.00/M input tokens
Starting at $0.00/M output tokens
No ratings yetSign in to rate

Select Provider

All Providers for Qwen-Audio-3.0-TTS Flash

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Alibaba Cloud
Context: 20k
Per Character Pricing
Input text$0.015/1K chars
Get Started