OpenTranscription
OpenTranscription
RankerModelsPlayground
OpenTranscription
OpenTranscription

One API to every speech-to-text model worth using. Compare them on your audio, route to the best one, pay per second.

Platform status

Product

RankerModelsTranscriptionsPlaygroundBlog

Developers

DocumentationReliabilityAPI VersioningStatus

Legal

Privacy PolicyTerms of ServiceSupport
© 2026 OpenTranscription
Back to Models

Chirp 3 vs Whisper Large V3

2 models compared on the same benchmark — accuracy, latency, price, language coverage and capabilities. Chirp 3 leads on raw accuracy; Whisper Large V3 is fastest; Whisper Large V3 covers the most languages; Whisper Large V3 is cheapest.

Open both in playground

Verdict · who wins what

Most accurate

Chirp 3

10.0% WER

Google Cloud

Fastest

Whisper Large V3

0.3× realtime

OpenAI

Most languages

Whisper Large V3

98 languages

OpenAI

Cheapest

Whisper Large V3

$0.0060/min

OpenAI

Comparing

2 models · same benchmark

1

Batch

Chirp 3

Google Cloud

$0.0107/min
#4 · 45.0

Google's latest generative ASR foundation model — 85+ languages

Use This ModelDetails →

2

Batch

Whisper Large V3

OpenAI

$0.0060/min
#24 · 69.3

OpenAI's Whisper large-v3 model

Use This ModelDetails →

Overview

30-day benchmark average

Overall score

0–100

45.0

69.3

Accuracy rank

of 35

#4

#24

Word error rate

WER

10.0%

17.7%

Character error rate

CER

5.8%

13.1%

Match error rate

MER

9.7%

16.6%

Word info lost

WIL

14.1%

21.6%

Languages

68

98

Cost vs. accuracy

overall WER vs price

Chirp 3

Whisper Large V3

field · 33 models

lower-left is better

Accuracy by category

WER · lower is better

code_switching

17.1%
62.7%

conversational

4.7%
11.2%

finance

7.1%
6.0%

general

8.2%
14.9%

legal

2.9%
2.4%

medical

1.9%
6.6%

noisy

29.2%
36.3%

technical

3.7%
6.5%

Accuracy by accent

WER · English

african

20.0%
24.6%

australian

16.9%
15.4%

indian

27.0%
19.1%

uk

3.0%
3.0%

us

5.7%
3.8%

Speed & latency

p50 → p99

Processing speed

× realtime

0.3×

0.3×

Latency p50

11.0s

3.5s

Latency p90

11.2s

4.3s

Latency p99

11.2s

4.3s

Realtime / streaming

Streaming benchmark averages

Realtime rank

—

—

Realtime score

0–100

—

—

Time to first word

—

—

Final drain

—

—

Real-time factor

—

—

Flicker

—

—

Cadence

—

—

Reliability

Rolling 30 days

Uptime

69.23%

100.00%

Error rate

30.77%

0.00%

Pricing

billed per second

Rate

per minute

$0.0107

$0.0060

Billing granularity

per 15s

per 1s

1 hour of audio

$0.64

$0.36

100 hours

$64.08

$36.00

1,000 hours

$640.80

$360.00

Capabilities

Feature support

Speaker diarization

Yes

No

Word-level timestamps

No

Yes

Code-switching

Yes

No

Auto-detect language

Yes

Yes

Live / streaming

No

No

Custom vocabulary

Yes

No

Coverage & deployment

Languages · formats · regions

Languages

68

98

Mode

Batch

Batch

Region

United States

United States

Audio formats

mp3
wav
flac
ogg
webm
amr
mp3
wav
flac
m4a
ogg
webm

Max file size

No limit

25 MB

Max duration

8 h

No limit

Compliance

HIPAA
SOC2
GDPR
SOC2
GDPR

Frequently asked questions

Is Chirp 3 or Whisper Large V3 more accurate?

Chirp 3 is more accurate, with a 10.0% word error rate versus 17.7% for Whisper Large V3, on our standardized benchmark.

Which is cheaper, Chirp 3 or Whisper Large V3?

Whisper Large V3 is cheaper at $0.0060/min versus $0.0107/min for Chirp 3.

Should you choose Chirp 3 or Whisper Large V3?

Choose Chirp 3 for accuracy; choose Whisper Large V3 for speed, lower cost, and language coverage.