Whisper Large V3 is a speech-to-text model from OpenAI. In our standardized benchmarks it reaches 17.67% word error rate, ranking #24 of 35 models tested. It supports 98 languages and runs at $0.006/min — mid-range on price.
Its nearest benchmarked alternative is Soniox STT (Async). Best suited for multilingual workloads and medical transcription.
Overall Score
69.3
#24 of 35
Word Error Rate
17.67%
Character Error Rate
13.06%
Match Error Rate
16.55%
Word Info Lost
21.63%
Avg Latency
3.3s
Benchmarks Run
35
last 30 runs
Word error rate
17.67%
↓ 26.6% · 30d
3 runs ago
latest
Avg latency
3.3s
↓ 23.7% · 30d
3 runs ago
latest
overall WER · 35 models
Whisper Large V3
field · 34 models
lower-left is better
8 categories
WER
lower is better
Legal
2.4%
Finance
6.0%
Technical
6.5%
Medical
6.6%
Conversational
11.2%
General
14.9%
Noisy Environment
36.3%
Code-Switching
62.7%
5 accents
Rate
$0.006
/min
Per-second billing. Bring your own provider key and pay your provider directly — a 5% routing fee applies (first 100 min/mo free).
Set up BYOKCost estimator
1 hour
$0.36
10 hours
$3.60
100 hours
$36
1,000 hours
$360
Billed per second of audio processed.
Model ID
openai/whisper-large-v3
Authenticate every request with your secret API key as a Bearer token. Issue a key from your dashboard.
Use it with your agent
Paste this into Claude Code, Cursor, Codex or any agent with a shell. It installs the CLI and the OpenTranscription skill, then transcribes with this model.
Quickstart
Upload your audio, create a job with this model, then poll the job or set a webhook_url to be notified when it completes.
Key parameters
file_path
string
required
Storage path returned by the upload step.
model
string
required
The model to run this job on.
models
string[]
Ordered fallback chain: primary first, then backups tried in order if a provider fails.
language
string
ISO 639-1 language code. Omit to auto-detect.
webhook_url
string
HTTPS URL notified with the result when the job completes. Delivered events are signed (X-OT-Signature).
title
string
Display name for this job in the dashboard. Falls back to the file name.
Response
A completed job returns the transcript in the OpenTranscription Unified Schema (OTUS).
Webhooks — skip polling
We POST a signed event to your webhook_url when a job completes or fails; verify the X-OT-Signature (HMAC-SHA256, reject if older than 300 s) and dedupe on the event id, then fetch the full transcript via GET /api/v1/transcriptions/{id}.
Rate-limited per tier — see the X-RateLimit-* response headers.
Full API referenceUptime · 30d
100.00%
Error rate
0.00%
Avg latency
33.7s
Last incident
None in 30d
from the same field
same benchmark, every match
Supported Languages
Supported Formats
Features
Limits
Max file size
25 MB
Max duration
No documented limit