Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Qwen · Voice model

Qwen-Audio-3.0-TTS Flash

qwen/qwen-audio-3.0-tts-flash
  • Voice
Released Jul 23, 2026

Qwen-Audio-3.0-TTS Flash is Alibaba's fast, cost-efficient text-to-speech model, generating spoken audio from text via the DashScope Speech Synthesizer API.

Specifications

Input
text, up to 20,000 characters per call
Output
mp3 (default) or raw pcm
Voice
provider voice ids via `voice`; omitted = the provider's default when it has one
Billing
per character of input text

Use it

POST /api/v1/audio/speech · MCP speak · full reference

curl https://omnirail.org/api/v1/audio/speech \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen-audio-3.0-tts-flash","input":"Hello from OmniRail. Every model, one key.","response_format":"mp3"}' \
  --output hello.mp3