Blog
Analysis, model profiles and deep dives on speech-to-text.
Analysis, model profiles and deep dives on speech-to-text.

Published July 3, 2026
A practitioner's breakdown of Deepgram Nova-3: WER claims, sub-300 ms streaming latency, pricing, languages, deployment options, and where it falls short.

Published July 3, 2026
What Scribe v2 actually changed from v1: features, pricing, benchmark results, API limits, and the architecture details ElevenLabs still won't publish.
Published July 3, 2026
Reference profile of ElevenLabs Scribe v2, a batch speech-to-text model released January 9, 2026: features, benchmarks, pricing, limits, and sources.
Published July 3, 2026
Reference profile of OpenAI's gpt-4o-transcribe speech-to-text model: release date, pricing, API features, benchmarks, and disclosed specifications.

Published July 3, 2026
A practitioner's look at gpt-4o-transcribe: pricing, API surface, benchmark evidence, and why OpenAI now recommends the mini model over it.

Published July 3, 2026
What Chirp 3 really is: Google's STT and TTS model family, its streaming limits, real pricing math, and how it compares to OpenAI, ElevenLabs, and Deepgram.
Published July 3, 2026
Reference profile of Google Cloud Chirp 3: multilingual speech-to-text in Speech-to-Text V2, Chirp 3 HD voices, pricing, limits, and benchmarks.
Published July 3, 2026
Reference profile of Google Cloud Speech-to-Text's default model, a general-purpose legacy baseline retained for backwards compatibility.
Published July 3, 2026
Reference profile of Google Cloud Speech-to-Text latest_long, a Conformer-based long-form transcription model: features, pricing, limits, history.
Published July 3, 2026
Reference profile of Google Cloud Speech-to-Text latest_short, a rolling Conformer-based model tag for short utterances and command-style speech.

Published July 3, 2026
What Google Cloud STT's default model actually is, why Google calls it legacy, and when it still beats routing audio to Chirp or the latest models.

Published July 3, 2026
A practitioner's guide to Google Cloud Speech-to-Text latest_long: Conformer roots, pricing, quotas, diarization, and how it compares to V2 and Chirp.