Blog
Analysis, model profiles and deep dives on speech-to-text.
Analysis, model profiles and deep dives on speech-to-text.
Published July 3, 2026
Reference profile of OpenAI's gpt-4o-transcribe speech-to-text model: release date, pricing, API features, benchmarks, and disclosed specifications.

Published July 3, 2026
A practitioner's look at gpt-4o-transcribe: pricing, API surface, benchmark evidence, and why OpenAI now recommends the mini model over it.

Published July 3, 2026
What Chirp 3 really is: Google's STT and TTS model family, its streaming limits, real pricing math, and how it compares to OpenAI, ElevenLabs, and Deepgram.
Published July 3, 2026
Reference profile of Google Cloud Chirp 3: multilingual speech-to-text in Speech-to-Text V2, Chirp 3 HD voices, pricing, limits, and benchmarks.
Published July 3, 2026
Reference profile of Google Cloud Speech-to-Text's default model, a general-purpose legacy baseline retained for backwards compatibility.
Published July 3, 2026
Reference profile of Google Cloud Speech-to-Text latest_long, a Conformer-based long-form transcription model: features, pricing, limits, history.
Published July 3, 2026
Reference profile of Google Cloud Speech-to-Text latest_short, a rolling Conformer-based model tag for short utterances and command-style speech.

Published July 3, 2026
What Google Cloud STT's default model actually is, why Google calls it legacy, and when it still beats routing audio to Chirp or the latest models.

Published July 3, 2026
A practitioner's guide to Google Cloud Speech-to-Text latest_long: Conformer roots, pricing, quotas, diarization, and how it compares to V2 and Chirp.

Published July 3, 2026
Why Google's latest_short model is built for short utterances, not short files, and when running it through batch recognition actually makes sense.
Published July 3, 2026
Reference profile of Google's command_and_search transcription model in Cloud Speech-to-Text, a legacy short-utterance model for voice commands and voice search.

Published July 3, 2026
The history, architecture, and current status of Google's command_and_search speech model, from 2016 Cloud Speech API beta to legacy status behind Chirp.