OpenTranscription
OpenTranscription
RankerModelsPlayground

Blog

Analysis, model profiles and deep dives on speech-to-text.

OpenTranscription
OpenTranscription

One API to every speech-to-text model worth using. Compare them on your audio, route to the best one, pay per second.

Platform status

Product

RankerModelsTranscriptionsPlaygroundBlog

Developers

DocumentationReliabilityAPI VersioningStatus

Legal

Privacy PolicyTerms of ServiceSupport
© 2026 OpenTranscription

Published July 3, 2026

Deepgram Nova-3: the enterprise ASR workhorse you can buy but not inspect

A practitioner's breakdown of Deepgram Nova-3: WER claims, sub-300 ms streaming latency, pricing, languages, deployment options, and where it falls short.

Read post

Published July 3, 2026

ElevenLabs Scribe v2: a top-tier transcription product built on an undisclosed model

What Scribe v2 actually changed from v1: features, pricing, benchmark results, API limits, and the architecture details ElevenLabs still won't publish.

Read post

Published July 3, 2026

ElevenLabs Scribe v2: model profile

Reference profile of ElevenLabs Scribe v2, a batch speech-to-text model released January 9, 2026: features, benchmarks, pricing, limits, and sources.

Read post

Published July 3, 2026

GPT-4o Transcribe: model profile

Reference profile of OpenAI's gpt-4o-transcribe speech-to-text model: release date, pricing, API features, benchmarks, and disclosed specifications.

Read post

Published July 3, 2026

GPT-4o Transcribe: what OpenAI ships, claims, and still won't tell you

A practitioner's look at gpt-4o-transcribe: pricing, API surface, benchmark evidence, and why OpenAI now recommends the mini model over it.

Read post

Published July 3, 2026

Google Cloud Chirp 3: capabilities, costs, and where it actually wins

What Chirp 3 really is: Google's STT and TTS model family, its streaming limits, real pricing math, and how it compares to OpenAI, ElevenLabs, and Deepgram.

Read post

Published July 3, 2026

Google Cloud Chirp 3: model profile

Reference profile of Google Cloud Chirp 3: multilingual speech-to-text in Speech-to-Text V2, Chirp 3 HD voices, pricing, limits, and benchmarks.

Read post

Published July 3, 2026

Google Cloud Speech-to-Text default: model profile

Reference profile of Google Cloud Speech-to-Text's default model, a general-purpose legacy baseline retained for backwards compatibility.

Read post

Published July 3, 2026

Google Cloud Speech-to-Text latest_long: model profile

Reference profile of Google Cloud Speech-to-Text latest_long, a Conformer-based long-form transcription model: features, pricing, limits, history.

Read post

Published July 3, 2026

Google Cloud latest_short: model profile

Reference profile of Google Cloud Speech-to-Text latest_short, a rolling Conformer-based model tag for short utterances and command-style speech.

Read post

Published July 3, 2026

Google Cloud's default speech model is legacy code that refuses to die

What Google Cloud STT's default model actually is, why Google calls it legacy, and when it still beats routing audio to Chirp or the latest models.

Read post

Published July 3, 2026

Google Cloud's latest_long model: what it is, what it costs, and when to pick something else

A practitioner's guide to Google Cloud Speech-to-Text latest_long: Conformer roots, pricing, quotas, diarization, and how it compares to V2 and Chirp.

Read post
Previous

Page 4 of 5

Next