Gladia Solaria-3 — EU-hosted batch speech-to-text tuned for the most accurate business audio across five core European languages (English, French, German, Spanish, Italian), with speaker diarization, custom vocabulary, and word-level timestamps.
5 languages · benchmarked
Overall Score
52.2
#3 of 35
Word Error Rate
9.80%
Character Error Rate
5.76%
Match Error Rate
9.63%
Word Info Lost
14.69%
Avg Latency
7.6s
Benchmarks Run
20
last 30 runs
Word error rate
9.80%
↓ 16.4% · 30d
3 runs ago
latest
Avg latency
7.6s
↑ 0.4% · 30d
3 runs ago
latest
overall WER · 35 models
Gladia Solaria-3
field · 34 models
lower-left is better
6 categories
WER
lower is better
Legal
2.9%
Conversational
7.0%
Medical
8.9%
General
9.0%
Finance
9.5%
Code-Switching
24.8%
5 accents
Rate
$0.0101
/min
Per-second billing. Bring your own provider key and pay your provider directly — a 5% routing fee applies (first 100 min/mo free).
Set up BYOKCost estimator
1 hour
$0.61
10 hours
$6.08
100 hours
$60.84
1,000 hours
$608.40
Billed per second of audio processed.
Model ID
gladia/solaria-3
Authenticate every request with your secret API key as a Bearer token. Issue a key from your dashboard.
Use it with your agent
Paste this into Claude Code, Cursor, Codex or any agent with a shell. It installs the CLI and the OpenTranscription skill, then transcribes with this model.
Quickstart
Upload your audio, create a job with this model, then poll the job or set a webhook_url to be notified when it completes.
Key parameters
file_path
string
required
Storage path returned by the upload step.
model
string
required
The model to run this job on.
models
string[]
Ordered fallback chain: primary first, then backups tried in order if a provider fails.
language
string
ISO 639-1 language code. Omit to auto-detect.
diarization
boolean
Label speakers (A, B, …) in the output.
custom_words
string[]
Names, jargon and product terms to bias the model toward. Up to 1000 entries.
vocabulary_list_id
uuid
A vocabulary list saved in your workspace. Merged with custom_words when both are sent.
webhook_url
string
HTTPS URL notified with the result when the job completes. Delivered events are signed (X-OT-Signature).
title
string
Display name for this job in the dashboard. Falls back to the file name.
Response
A completed job returns the transcript in the OpenTranscription Unified Schema (OTUS).
Webhooks — skip polling
We POST a signed event to your webhook_url when a job completes or fails; verify the X-OT-Signature (HMAC-SHA256, reject if older than 300 s) and dedupe on the event id, then fetch the full transcript via GET /api/v1/transcriptions/{id}.
Rate-limited per tier — see the X-RateLimit-* response headers.
Full API referenceSupported Languages
Supported Formats
Features
Limits
Max file size
1000 MB
Max duration
2 h 15 min
Gladia Solaria-3 is a speech-to-text model from Gladia. In our standardized benchmarks it reaches 9.80% word error rate, ranking #3 of 35 models tested. It supports 5 languages and runs at $0.01/min — a premium option.
Its nearest benchmarked alternative is Soniox STT (Async). Best suited for multi-speaker conversations, accuracy-critical work, and legal transcription.