OpenTranscription
OpenTranscription
RankerModelsPlayground

Blog

Analysis, model profiles and deep dives on speech-to-text.

OpenTranscription
OpenTranscription

One API to every speech-to-text model worth using. Compare them on your audio, route to the best one, pay per second.

Platform status

Product

RankerModelsTranscriptionsPlaygroundBlog

Developers

DocumentationReliabilityAPI VersioningStatus

Legal

Privacy PolicyTerms of ServiceSupport
© 2026 OpenTranscription

Published July 31, 2026

Speaker Diarization: A Developer's Technical Guide

Discover how speaker diarization enhances meeting transcriptions, call-center analytics, and media indexing. Learn valuable developer insights!

Read post

Published July 30, 2026

Google Meet Transcription for Developers: Batch and Real-Time Patterns

Discover the best approaches for google meet transcription, including batch and real-time patterns, to enhance data processing and compliance.

Read post

Published July 3, 2026

Amazon Transcribe Medical: what AWS actually ships, and what it won't tell you

What Amazon Transcribe Medical offers in 2026: features, specs, pricing vs Google and Nuance, HIPAA posture, research clues, and where it falls short.

Read post

Published July 3, 2026

Chirp 3: inside Google Cloud's 2025 speech stack, from HD voices to transcription

What Google Cloud Chirp 3 actually is: release timeline, WER and Elo benchmarks, pricing, specs, known issues, and how it compares to Azure and ElevenLabs.

Read post

Published July 3, 2026

Deepgram Base in 2026: what the legacy model still does well

Where Deepgram Base fits in 2026: API behavior, variants, latency, specs, limitations, and when to pick Nova-3 or Flux instead.

Read post

Published July 3, 2026

Deepgram Enhanced behind Future AGI's Agent Command Center: what the public record actually shows

A close read of Deepgram's Enhanced STT tier and Future AGI's Agent Command Center gateway, with full specs and the documentation gap between the two.

Read post

Published July 3, 2026

Deepgram Flux: turn detection moves inside the speech model

What Deepgram Flux actually is: a turn-aware streaming STT model for voice agents, its event API, specs, pricing, benchmarks, and where it falls short.

Read post

Published July 3, 2026

Deepgram Nova-3: the enterprise ASR workhorse you can buy but not inspect

A practitioner's breakdown of Deepgram Nova-3: WER claims, latency, pricing, languages, deployment, specs, limitations, and where it falls short.

Read post

Published July 3, 2026

ElevenLabs Scribe v2: a top-tier transcription product built on an undisclosed model

What Scribe v2 actually changed from v1: features, pricing, benchmark results, API limits, and the architecture details ElevenLabs still won't publish.

Read post

Published July 3, 2026

GPT-4o Transcribe: what OpenAI ships, claims, and still won't tell you

GPT-4o Transcribe: OpenAI's pricing, API surface, specs, benchmark evidence, and known limitations — including why OpenAI now recommends the mini model.

Read post

Published July 3, 2026

Google Cloud Chirp 3: capabilities, costs, and where it actually wins

Google Cloud Chirp 3's real capabilities, streaming limits, pricing math, and how it compares to OpenAI, ElevenLabs, and Deepgram for speech-to-text and TTS.

Read post

Published July 3, 2026

Google Cloud's default speech model is legacy code that refuses to die

What Google Cloud STT's default model actually is, why Google calls it legacy, and when it still beats routing audio to Chirp or the latest models.

Read post
Previous
12345
Next