Optimized for human-to-bot interactions (IVR, voice assistants)
1 languages · benchmarked
Nova-2 Conversational AI is a speech-to-text model from Deepgram. In our standardized benchmarks it reaches 18.23% word error rate, ranking #26 of 35 models tested. It supports 1 language and runs at $0.006/min — mid-range on price.
Its nearest benchmarked alternative is Soniox STT (Async). Best suited for live captioning and streaming transcription and multi-speaker conversations.
Overall Score
67.6
#26 of 35
Word Error Rate
18.23%
Character Error Rate
13.27%
Match Error Rate
17.68%
Word Info Lost
24.14%
Avg Latency
3.9s
Benchmarks Run
25
last 30 runs
Word error rate
18.23%
↓ 4.6% · 30d
3 runs ago
latest
Avg latency
3.9s
↑ 76.5% · 30d
3 runs ago
latest
overall WER · 35 models
Nova-2 Conversational AI
field · 34 models
lower-left is better
8 categories
WER
lower is better
Legal
5.0%
Technical
5.2%
Medical
7.3%
Conversational
10.5%
Finance
11.3%
General
11.6%
Noisy Environment
34.9%
Code-Switching
66.9%
| Category | WER | CER | MER | WIL | Latency | Benchmarks |
|---|---|---|---|---|---|---|
Code-Switching | 66.90% | 63.17% | 66.07% | 70.64% | 1.9s | 2 |
Conversational | 10.49% | 7.04% | 10.44% | 12.91% | 4.7s | 2 |
Finance | 11.29% | 5.45% | 11.21% | 18.73% | 3.6s | 2 |
General | 11.57% | 5.94% | 11.47% | 18.70% | 4.4s | 9 |
Legal | 4.98% | 1.96% | 4.94% | 7.30% | 4.1s | 2 |
Medical | 7.28% | 2.78% | 6.98% | 9.80% | 5.0s | 2 |
Noisy Environment | 34.86% | 27.84% | 32.36% | 45.50% | 3.3s | 4 |
Technical | 5.17% | 3.11% | 5.09% | 7.21% | 3.2s | 2 |
5 accents
| Accent | WER | Benchmarks |
|---|---|---|
African | 27.69% | 1 |
Australian | 21.54% | 1 |
Indian | 26.98% | 1 |
British | 8.96% | 1 |
American | 13.21% | 1 |
Realtime Score
77.4
TTFW (P50)
1013 ms
Flicker
10.5%
Cadence
0.8/s
RTF
1.01×
Word Error Rate
18.80%
Character Error Rate
12.09%
Benchmarks Run
25
Rate
$0.0058
/min
Per-second billing. Bring your own provider key and pay your provider directly — a 5% routing fee applies (first 100 min/mo free).
Set up BYOKCost estimator
1 hour
$0.35
10 hours
$3.49
100 hours
$34.92
1,000 hours
$349.20
Billed per second of audio processed.
Model ID
deepgram/nova-2-conversationalai
Authenticate every request with your secret API key as a Bearer token. Issue a key from your dashboard.
This model also supports realtime streaming over WebSocket.
Use it with your agent
Paste this into Claude Code, Cursor, Codex or any agent with a shell. It installs the CLI and the OpenTranscription skill, then transcribes with this model.
Quickstart
Upload your audio, create a job with this model, then poll the job or set a webhook_url to be notified when it completes.
Key parameters
file_path
string
required
Storage path returned by the upload step.
model
string
required
The model to run this job on.
models
string[]
Ordered fallback chain: primary first, then backups tried in order if a provider fails.
language
string
ISO 639-1 language code. Omit to auto-detect.
diarization
boolean
Label speakers (A, B, …) in the output.
word_timestamps
boolean
Per-word start, end and confidence. Defaults to true; send false to leave them out.
custom_words
string[]
Names, jargon and product terms to bias the model toward. Up to 1000 entries.
vocabulary_list_id
uuid
A vocabulary list saved in your workspace. Merged with custom_words when both are sent.
webhook_url
string
HTTPS URL notified with the result when the job completes. Delivered events are signed (X-OT-Signature).
title
string
Display name for this job in the dashboard. Falls back to the file name.
Response
A completed job returns the transcript in the OpenTranscription Unified Schema (OTUS).
Webhooks — skip polling
We POST a signed event to your webhook_url when a job completes or fails; verify the X-OT-Signature (HMAC-SHA256, reject if older than 300 s) and dedupe on the event id, then fetch the full transcript via GET /api/v1/transcriptions/{id}.
Rate-limited per tier — see the X-RateLimit-* response headers.
Full API referenceSupported Languages
Supported Formats
Features
Limits
Max file size
2 GB
Max duration
No documented limit