Catalog · 43 models · 15 providers
Every speech-to-text model worth using, one rate card.
43
Total models
105
Languages
$0.0017
Cheapest /min
—
Best WER
2 models
filtered from 43
Soniox
Soniox async batch transcription — one multilingual model across 60+ languages with speaker diarization, per-word timestamps, and automatic language identification, at a low flat token-based rate.
Speed Factor
0.1x
Billing
per 1s
Retention
25 languages
(view all)
Soniox
Soniox realtime streaming transcription — low-latency multilingual STT across 60+ languages with speaker diarization, per-word timestamps, and automatic language identification over a single WebSocket.
Realtime model — batch benchmarks not applicable
Speed Factor
0.1x
Billing
per 1s
Retention
25 languages
(view all)