OpenTranscription
OpenTranscription
RankerModelsPlayground
OpenTranscription
OpenTranscription

One API to every speech-to-text model worth using. Compare them on your audio, route to the best one, pay per second.

Platform status

Product

RankerModelsTranscriptionsPlaygroundBlog

Developers

DocumentationReliabilityAPI VersioningStatus

Legal

Privacy PolicyTerms of ServiceSupport
© 2026 OpenTranscription
Back to Models

Scribe v2 vs Gladia Solaria-3

2 models compared on the same benchmark — accuracy, latency, price, language coverage and capabilities. Scribe v2 leads on raw accuracy; Scribe v2 covers the most languages; Scribe v2 is cheapest.

Open both in playground

Verdict · who wins what

Most accurate

Scribe v2

9.2% WER

ElevenLabs

Fastest

—

—

Most languages

Scribe v2

76 languages

ElevenLabs

Cheapest

Scribe v2

$0.0040/min

ElevenLabs

Comparing

2 models · same benchmark

1

Batch

Scribe v2

ElevenLabs

$0.0040/min
#2 · 78.2

State-of-the-art batch STT — 90+ languages, speaker diarization, audio tagging

Use This ModelDetails →

2

Batch

Gladia Solaria-3

Gladia

$0.0101/min
#3 · 52.2

Gladia Solaria-3 — EU-hosted batch speech-to-text tuned for the most accurate business audio across five core European languages (English, French, German, Spanish, Italian), with speaker diarization, custom vocabulary, and word-level timestamps.

Use This ModelDetails →

Overview

30-day benchmark average

Overall score

0–100

78.2

52.2

Accuracy rank

of 35

#2

#3

Word error rate

WER

9.2%

9.8%

Character error rate

CER

7.2%

5.8%

Match error rate

MER

8.5%

9.6%

Word info lost

WIL

12.3%

14.7%

Languages

76

5

Cost vs. accuracy

overall WER vs price

Scribe v2

Gladia Solaria-3

field · 33 models

lower-left is better

Accuracy by category

WER · lower is better

code_switching

14.6%
24.8%

conversational

7.6%
7.0%

finance

7.6%

Accuracy by accent

WER · English

african

21.5%
24.6%

australian

3.1%
13.9%

indian

9.5%

Speed & latency

p50 → p99

Processing speed

× realtime

0.2×

0.2×

Latency p50

3.4s

7.6s

Latency p90

4.0s

7.7s

Realtime / streaming

Streaming benchmark averages

Realtime rank

—

—

Realtime score

0–100

—

—

Time to first word

—

—

Final drain

—

—

Real-time factor

—

—

Flicker

—

—

Cadence

—

—

Reliability

Rolling 30 days

Uptime

100.00%

100.00%

Error rate

0.00%

0.00%

Pricing

billed per second

Rate

per minute

$0.0040

$0.0101

Billing granularity

per 1s

per 1s

1 hour of audio

$0.24

$0.61

100 hours

Capabilities

Feature support

Speaker diarization

Yes

Yes

Word-level timestamps

Yes

Yes

Code-switching

Yes

Coverage & deployment

Languages · formats · regions

Languages

76

5

Mode

Batch

Batch

Region

United States

EU (France)

Audio formats

9.5%

general

5.4%
9.0%

legal

2.8%
2.9%

medical

5.8%
8.9%

noisy

33.8%

—

technical

3.1%

—

25.4%

uk

0.0%
7.5%

us

0.0%
13.2%

Latency p99

4.0s

7.7s

$24.12

$60.84

1,000 hours

$241.20

$608.40

No

Auto-detect language

Yes

Yes

Live / streaming

No

No

Custom vocabulary

No

Yes

mp3
wav
flac
m4a
ogg
aac
mp3
wav
flac
m4a
ogg
webm
aac

Max file size

3 GB

1000 MB

Max duration

10 h

2 h 15 min

Compliance

SOC2
GDPR
GDPR

Frequently asked questions

Is Scribe v2 or Gladia Solaria-3 more accurate?

Scribe v2 is more accurate, with a 9.2% word error rate versus 9.8% for Gladia Solaria-3, on our standardized benchmark.

Which is cheaper, Scribe v2 or Gladia Solaria-3?

Scribe v2 is cheaper at $0.0040/min versus $0.0101/min for Gladia Solaria-3.

Should you choose Scribe v2 or Gladia Solaria-3?

Scribe v2 is the stronger choice overall, leading on accuracy, lower cost, and language coverage.