Qwen3 TTS Flash API

alibaba/qwen3-tts-flash
Qwen3-TTS-Flash is a fast, high-quality text-to-speech model optimized for natural and expressive multilingual voice synthesis with ultra-low latency.
Output
$0.013 / 1M tokens
Released
Jan 5, 2026

How to use Qwen3 TTS Flash API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to alibaba/qwen3-tts-flash.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/tts",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "alibaba/qwen3-tts-flash",
      "text": "Hello from AI/ML API"
    },
)
print(r.json()["audio"])  # URL to the generated audio
const r = await fetch("https://api.aimlapi.com/v1/tts", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "alibaba/qwen3-tts-flash",
    "text": "Hello from AI/ML API"
  }),
});
const { audio } = await r.json();
console.log(audio); // URL to the generated audio
curl -X POST https://api.aimlapi.com/v1/tts \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"alibaba/qwen3-tts-flash","text":"Hello from AI/ML API"}'

# the JSON response contains "audio" — a URL to the generated speech

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Qwen3 TTS Flash API Pricing

TypePrice
Input
$13 / 1M characters
Output
$0.013 / 1M tokens

Qwen3 TTS Flash Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
TTS Arena
926 (Elo)
Human preference Elo from blind pairwise listening comparisons of speech samples, measured independently by Artificial AnalysisSourceJuly 16, 2026

Qwen3 TTS Flash vs other models

ModelInputOutputContextBest for
$13 / 1M characters
$0.013 / 1M tokens
Speech synthesis
$130 / 1M characters
$130 / 1M
Speech synthesis
$130 / 1M characters
$0.13 / 1M tokens
Speech synthesis
$78 / 1M characters
$0.078 / 1M tokens
Speech synthesis

Frequently asked questions

Qwen3 TTS Flash takes text as input and returns audio.

Qwen3 TTS Flash became available on January 5, 2026.

Qwen3 TTS Flash is priced at $0.013 / 1M tokens.

Qwen3 TTS Flash is billed per generation — a fixed charge per output rather than by prompt length.

Qwen3 TTS Flash was built by Alibaba Cloud.

Send a request to https://api.aimlapi.com/v1/tts with alibaba/qwen3-tts-flash as the model id.

Yes. Qwen3 TTS Flash is served through AI/ML API, so the same key and endpoint format used for other models applies.

It supports multiple languages and dialects, including Chinese.

Qwen3 TTS Flash offers 17 voices.

Start building with Qwen3 TTS Flash

Get API Key
1000+ models, one API.