OpenTranscription
OpenTranscription
RankerModelsPlayground
OpenTranscription
OpenTranscription

One API to every speech-to-text model worth using. Compare them on your audio, route to the best one, pay per second.

Platform status

Product

RankerModelsTranscriptionsPlaygroundBlog

Developers

DocumentationReliabilityAPI VersioningStatus

Legal

Privacy PolicyTerms of ServiceSupport
© 2026 OpenTranscription
Back to Models

Nova-2 vs Whisper Large V3

2 models compared on the same benchmark — accuracy, latency, price, language coverage and capabilities. Whisper Large V3 leads on raw accuracy; Nova-2 is fastest; Nova-2 covers the most languages; Whisper Large V3 is cheapest.

Open both in playground

Verdict · who wins what

Most accurate

Whisper Large V3

14.4% WER

Groq

Fastest

Nova-2

0.1× realtime

Deepgram

Most languages

Nova-2

33 languages

Deepgram

Cheapest

Whisper Large V3

$0.0019/min

Groq

Comparing

2 models · same benchmark

1

Batch & Realtime

Nova-2

Deepgram

$0.0058/min
#18 · 76.2

Deepgram's Nova-2 speech recognition

Use This ModelDetails →

2

Batch

Whisper Large V3

Groq

$0.0019/min
#11 · 79.9

OpenAI Whisper large-v3 served on Groq's LPU hardware — top Whisper accuracy at Groq speed and cost, with word-level timestamps and automatic language detection.

Use This ModelDetails →

Overview

30-day benchmark average

Overall score

0–100

76.2

79.9

Accuracy rank

of 35

#18

#11

Word error rate

WER

16.5%

14.4%

Character error rate

CER

11.9%

11.2%

Match error rate

MER

15.4%

13.3%

Word info lost

WIL

21.0%

18.2%

Languages

33

12

Cost vs. accuracy

overall WER vs price

Nova-2

Whisper Large V3

field · 33 models

lower-left is better

Accuracy by category

WER · lower is better

code_switching

63.7%
68.5%

conversational

10.2%
10.9%

finance

11.3%
6.4%

Accuracy by accent

WER · English

african

26.2%
26.2%

australian

15.4%
12.3%

indian

25.4%
19.1%

Speed & latency

p50 → p99

Processing speed

× realtime

0.1×

0.0×

Latency p50

1.6s

3.1s

Latency p90

3.7s

3.6s

Realtime / streaming

Streaming benchmark averages

Realtime rank

of 17

#4

—

Realtime score

0–100

78.0

—

Time to first word

986ms

—

Reliability

Rolling 30 days

Uptime

83.33%

—

Error rate

16.67%

—

Pricing

billed per second

Rate

per minute

$0.0058

$0.0019

Billing granularity

per 1s

per 1s

1 hour of audio

$0.35

$0.11

100 hours

$34.92

Capabilities

Feature support

Speaker diarization

Yes

No

Word-level timestamps

Yes

Yes

Code-switching

Yes

Coverage & deployment

Languages · formats · regions

Languages

33

12

Mode

Batch & Realtime

Batch

Region

United States

United States

Audio formats

general

13.1%
8.1%

legal

3.8%
2.7%

medical

7.3%
2.8%

noisy

31.3%
36.3%

technical

4.9%
5.3%

uk

11.9%
6.0%

us

9.4%
5.7%

Latency p99

3.7s

3.6s

Final drain

192ms

—

Real-time factor

1.02×

—

Flicker

10.5%

—

Cadence

0.8/s

—

$11.16

1,000 hours

$349.20

$111.60

No

Auto-detect language

Yes

Yes

Live / streaming

Yes

No

Custom vocabulary

Yes

No

mp3
wav
flac
m4a
ogg
webm
mp3
wav
flac
m4a
ogg
webm

Max file size

2 GB

100 MB

Max duration

No limit

No limit

Compliance

HIPAA
SOC2
GDPR

—

Frequently asked questions

Is Nova-2 or Whisper Large V3 more accurate?

Whisper Large V3 is more accurate, with a 14.4% word error rate versus 16.5% for Nova-2, on our standardized benchmark.

Which is cheaper, Nova-2 or Whisper Large V3?

Whisper Large V3 is cheaper at $0.0019/min versus $0.0058/min for Nova-2.

Should you choose Nova-2 or Whisper Large V3?

Choose Nova-2 for speed and language coverage; choose Whisper Large V3 for accuracy and lower cost.