OpenTranscription
OpenTranscription
RankerModelsPlayground

Blog

Analysis, model profiles and deep dives on speech-to-text.

OpenTranscription
OpenTranscription

One API to every speech-to-text model worth using. Compare them on your audio, route to the best one, pay per second.

Platform status

Product

RankerModelsTranscriptionsPlaygroundBlog

Developers

DocumentationReliabilityAPI VersioningStatus

Legal

Privacy PolicyTerms of ServiceSupport
© 2026 OpenTranscription

Published July 3, 2026

Google Cloud's latest_long model: what it is, what it costs, and when to pick something else

A practitioner's guide to Google Cloud Speech-to-Text latest_long: Conformer roots, pricing, quotas, diarization, and how it compares to V2 and Chirp.

Read post

Published July 3, 2026

Google Cloud's latest_short and the batch paradox

Why Google's latest_short model is built for short utterances, not short files, and when running it through batch recognition actually makes sense.

Read post

Published July 3, 2026

Google's command_and_search model: the voice-search engine that quietly became legacy

The history, architecture, and current status of Google's command_and_search speech model, from 2016 Cloud Speech API beta to legacy status behind Chirp.

Read post

Published July 3, 2026

Ink-Whisper: how Cartesia rebuilt Whisper for real-time voice agents

What Cartesia's Ink-Whisper got right on latency, where its accuracy fell behind by 2026, and why it mattered more as a stepping stone than a benchmark.

Read post

Published July 3, 2026

Scribe v2 Realtime: ElevenLabs makes its play for live speech-to-text

Scribe v2 Realtime by ElevenLabs: specs, pricing, benchmarks, and known limitations behind the sub-150ms, 93.5%-accuracy, 90+ language claims.

Read post

Published July 3, 2026

Universal-3 Pro: what AssemblyAI shipped, and what it still won't say

AssemblyAI's Universal-3 Pro reviewed: promptable transcription, WER benchmarks, pricing, compliance caveats, and what the public record still hides.

Read post

Published July 3, 2026

Whisper large-v3 and the shift from open research to transcription infrastructure

How Whisper large-v3 went from OpenAI's MIT-licensed research release to the baseline of a managed transcription stack — specs, data, and what's unresolved.

Read post

Published July 3, 2026

Whisper on Azure: what Microsoft actually sells, and where it fits now

How Microsoft packages OpenAI's Whisper across Azure OpenAI and Azure Speech: limits, pricing signals, benchmarks, security, and where it fits in 2026.

Read post
Previous
12345