Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

Mistral · Transcription model

Voxtral Small 24B 2507 STT

mistralai/voxtral-small-24b-2507-stt
  • Transcription
Released Aug 13, 2026

Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.

Specifications

Input
wav, mp3, flac, m4a, ogg, webm, aac, opus · base64 JSON, multipart file or URL · up to 25 MB
Output
text; verbose_json adds language, duration, segments and word timestamps
Billing
per second of audio

Use it

POST /api/v1/audio/transcriptions · MCP transcribe · full reference

curl https://omnirail.org/api/v1/audio/transcriptions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" \
  -F model=mistralai/voxtral-small-24b-2507-stt \
  -F file=@meeting.mp3