OpenTranscription
OpenTranscription
RankerModelsPlayground
OpenTranscription
OpenTranscription

One API to every speech-to-text model worth using. Compare them on your audio, route to the best one, pay per second.

Platform status

Product

RankerModelsTranscriptionsPlaygroundBlog

Developers

DocumentationReliabilityAPI VersioningStatus

Legal

Privacy PolicyTerms of ServiceSupport
© 2026 OpenTranscription
Back to Models

Amazon Transcribe vs Whisper Large V3

2 models compared on the same benchmark — accuracy, latency, price, language coverage and capabilities. Amazon Transcribe leads on raw accuracy; Amazon Transcribe is fastest; Whisper Large V3 covers the most languages.

Open both in playground

Verdict · who wins what

Most accurate

Amazon Transcribe

13.0% WER

Amazon Web Services

Fastest

Amazon Transcribe

0.4× realtime

Amazon Web Services

Most languages

Whisper Large V3

98 languages

OpenAI

Cheapest

—

—

Comparing

2 models · same benchmark

1

Batch

Amazon Transcribe

Amazon Web Services

$0.0060/min
#8 · 51.5

AWS foundation model-powered ASR — 100+ languages

Use This ModelDetails →

2

Batch

Whisper Large V3

OpenAI

$0.0060/min
#24 · 69.3

OpenAI's Whisper large-v3 model

Use This ModelDetails →

Overview

30-day benchmark average

Overall score

0–100

51.5

69.3

Accuracy rank

of 35

#8

#24

Word error rate

WER

13.0%

17.7%

Character error rate

CER

9.2%

13.1%

Match error rate

MER

12.1%

16.6%

Word info lost

WIL

16.7%

21.6%

Languages

77

98

Cost vs. accuracy

overall WER vs price

Amazon Transcribe

Whisper Large V3

field · 33 models

lower-left is better

Accuracy by category

WER · lower is better

code_switching

55.8%
62.7%

conversational

6.9%
11.2%

finance

6.3%
6.0%

general

6.6%
14.9%

legal

6.2%
2.4%

medical

3.7%
6.6%

noisy

41.7%
36.3%

technical

3.2%
6.5%

Accuracy by accent

WER · English

african

24.6%
24.6%

australian

6.2%
15.4%

indian

3.2%
19.1%

uk

0.0%
3.0%

us

1.9%
3.8%

Speed & latency

p50 → p99

Processing speed

× realtime

0.4×

0.3×

Latency p50

12.4s

3.5s

Latency p90

13.0s

4.3s

Latency p99

13.0s

4.3s

Realtime / streaming

Streaming benchmark averages

Realtime rank

—

—

Realtime score

0–100

—

—

Time to first word

—

—

Final drain

—

—

Real-time factor

—

—

Flicker

—

—

Cadence

—

—

Reliability

Rolling 30 days

Uptime

100.00%

100.00%

Error rate

0.00%

0.00%

Pricing

billed per second

Rate

per minute

$0.0060

$0.0060

Billing granularity

per 1s

per 1s

1 hour of audio

$0.36

$0.36

100 hours

$36.00

$36.00

1,000 hours

$360.00

$360.00

Capabilities

Feature support

Speaker diarization

Yes

No

Word-level timestamps

Yes

Yes

Code-switching

Yes

No

Auto-detect language

Yes

Yes

Live / streaming

No

No

Custom vocabulary

No

No

Coverage & deployment

Languages · formats · regions

Languages

77

98

Mode

Batch

Batch

Region

United States

United States

Audio formats

mp3
wav
flac
m4a
ogg
webm
amr
mp3
wav
flac
m4a
ogg
webm

Max file size

2 GB

25 MB

Max duration

8 h

No limit

Compliance

HIPAA
SOC2
GDPR
SOC2
GDPR

Frequently asked questions

Is Amazon Transcribe or Whisper Large V3 more accurate?

Amazon Transcribe is more accurate, with a 13.0% word error rate versus 17.7% for Whisper Large V3, on our standardized benchmark.

Which is cheaper, Amazon Transcribe or Whisper Large V3?

Amazon Transcribe and Whisper Large V3 cost the same at $0.0060/min.

Should you choose Amazon Transcribe or Whisper Large V3?

Choose Amazon Transcribe for accuracy and speed; choose Whisper Large V3 for language coverage.