Qwen-Audio-3.0-TTS Flash

Alibaba's low-latency Qwen-Audio-3.0 text-to-speech model optimized for real-time interaction across 16 languages. Generates speech via the /v1/audio/speech endpoint.

qwen-audio-3.0-tts-flash
STABLEGet StartedView uptime
20,000 context
Released July 20, 2026
Starting at $0.00/M input tokens
Starting at $0.00/M output tokens
No ratings yetSign in to rate

Select Provider

All Providers for Qwen-Audio-3.0-TTS Flash

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Alibaba Cloud
Context: 20k
Per Character Pricing
Input text$0.015/1K chars
Get Started

Frequently asked questions

What is Qwen-Audio-3.0-TTS Flash?

Alibaba's low-latency Qwen-Audio-3.0 text-to-speech model optimized for real-time interaction across 16 languages. Generates speech via the /v1/audio/speech endpoint. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does Qwen-Audio-3.0-TTS Flash cost?

Pricing for Qwen-Audio-3.0-TTS Flash on LLM Gateway starts at $0.00 per million input tokens and $0.00 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of Qwen-Audio-3.0-TTS Flash?

Qwen-Audio-3.0-TTS Flash supports a context window of up to 20,000 tokens on its largest provider deployment.

Which providers serve Qwen-Audio-3.0-TTS Flash?

Qwen-Audio-3.0-TTS Flash is served by Alibaba Cloud through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does Qwen-Audio-3.0-TTS Flash support tool calling and structured outputs?

No. Qwen-Audio-3.0-TTS Flash does not currently support tool calling or structured JSON outputs through LLM Gateway.

When was Qwen-Audio-3.0-TTS Flash released?

Qwen-Audio-3.0-TTS Flash was released on July 20, 2026.