Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Qwen · Transcription model

Qwen3 ASR 1.7B

qwen/qwen3-asr-1.7b
  • Transcription
Released Aug 13, 2026

Qwen3 ASR 1.7B is an automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline inference...

Specifications

Input
wav, mp3, flac, m4a, ogg, webm, aac, opus · base64 JSON, multipart file or URL · up to 25 MB
Output
text; verbose_json adds language, duration, segments and word timestamps
Billing
per second of audio

Use it

POST /api/v1/audio/transcriptions · MCP transcribe · full reference

curl https://omnirail.org/api/v1/audio/transcriptions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" \
  -F model=qwen/qwen3-asr-1.7b \
  -F file=@meeting.mp3