OpenTranscription
OpenTranscription
RankerModelsPlayground
OpenTranscription
OpenTranscription

One API to every speech-to-text model worth using. Compare them on your audio, route to the best one, pay per second.

Platform status

Product

RankerModelsTranscriptionsPlaygroundBlog

Developers

DocumentationReliabilityAPI VersioningStatus

Legal

Privacy PolicyTerms of ServiceSupport
© 2026 OpenTranscription
Back to Models

Chirp 3 vs Whisper Large V3 Turbo

2 models compared on the same benchmark — accuracy, latency, price, language coverage and capabilities. Chirp 3 leads on raw accuracy; Chirp 3 is fastest; Chirp 3 covers the most languages; Whisper Large V3 Turbo is cheapest.

Open both in playground

Verdict · who wins what

Most accurate

Chirp 3

10.0% WER

Google Cloud

Fastest

Chirp 3

0.3× realtime

Google Cloud

Most languages

Chirp 3

68 languages

Google Cloud

Cheapest

Whisper Large V3 Turbo

$0.0007/min

Groq

Comparing

2 models · same benchmark

1

Batch

Chirp 3

Google Cloud

$0.0107/min
#4 · 45.0

Google's latest generative ASR foundation model — 85+ languages

Use This ModelDetails →

2

Batch

Whisper Large V3 Turbo

Groq

$0.0007/min
#10 · 81.3

OpenAI Whisper large-v3-turbo served on Groq's LPU hardware — very fast, very low cost batch transcription with word-level timestamps and automatic language detection.

Use This ModelDetails →

Overview

30-day benchmark average

Overall score

0–100

45.0

81.3

Accuracy rank

of 35

#4

#10

Word error rate

WER

10.0%

14.3%

Character error rate

CER

5.8%

10.6%

Match error rate

MER

9.7%

13.4%

Word info lost

WIL

14.1%

19.0%

Languages

68

12

Cost vs. accuracy

overall WER vs price

Chirp 3

Whisper Large V3 Turbo

field · 33 models

lower-left is better

Accuracy by category

WER · lower is better

code_switching

17.1%
49.9%

conversational

4.7%
13.7%

finance

7.1%

Accuracy by accent

WER · English

african

20.0%
24.6%

australian

16.9%
16.9%

indian

27.0%

Speed & latency

p50 → p99

Processing speed

× realtime

0.3×

0.0×

Latency p50

11.0s

2.6s

Latency p90

11.2s

3.4s

Realtime / streaming

Streaming benchmark averages

Realtime rank

—

—

Realtime score

0–100

—

—

Time to first word

—

—

Final drain

—

—

Real-time factor

—

—

Flicker

—

—

Cadence

—

—

Reliability

Rolling 30 days

Uptime

88.89%

100.00%

Error rate

11.11%

0.00%

Pricing

billed per second

Rate

per minute

$0.0107

$0.0007

Billing granularity

per 15s

per 1s

1 hour of audio

$0.64

$0.04

100 hours

$64.08

Capabilities

Feature support

Speaker diarization

Yes

No

Word-level timestamps

No

Yes

Code-switching

Yes

Coverage & deployment

Languages · formats · regions

Languages

68

12

Mode

Batch

Batch

Region

United States

United States

Audio formats

7.1%

general

8.2%
9.2%

legal

2.9%
3.8%

medical

1.9%
1.8%

noisy

29.2%
36.7%

technical

3.7%
8.3%
27.0%

uk

3.0%
6.0%

us

5.7%
9.4%

Latency p99

11.2s

3.4s

$3.96

1,000 hours

$640.80

$39.60

No

Auto-detect language

Yes

Yes

Live / streaming

No

No

Custom vocabulary

Yes

No

mp3
wav
flac
ogg
webm
amr
mp3
wav
flac
m4a
ogg
webm

Max file size

No limit

100 MB

Max duration

8 h

No limit

Compliance

HIPAA
SOC2
GDPR

—

Frequently asked questions

Is Chirp 3 or Whisper Large V3 Turbo more accurate?

Chirp 3 is more accurate, with a 10.0% word error rate versus 14.3% for Whisper Large V3 Turbo, on our standardized benchmark.

Which is cheaper, Chirp 3 or Whisper Large V3 Turbo?

Whisper Large V3 Turbo is cheaper at $0.0007/min versus $0.0107/min for Chirp 3.

Should you choose Chirp 3 or Whisper Large V3 Turbo?

Choose Chirp 3 for accuracy, speed, and language coverage; choose Whisper Large V3 Turbo for lower cost.