Blog
Analysis, model profiles and deep dives on speech-to-text.
Analysis, model profiles and deep dives on speech-to-text.

Published July 3, 2026
What Google Cloud Chirp 3 actually is: release timeline, WER and Elo benchmarks, pricing, known issues, and how it stacks up against Azure and ElevenLabs.
Published July 3, 2026
Reference profile of Google Cloud Chirp 3, a managed speech model family covering multilingual transcription, HD text-to-speech, and instant custom voice.

Published July 3, 2026
Where Deepgram Base fits in 2026: API behavior, variants, latency, concurrency, missing benchmarks, and when to pick Nova-3 or Flux instead.
Published July 3, 2026
Reference profile of Deepgram Base, a legacy speech-to-text model family with task-specific variants, batch and streaming APIs, and self-hosted deployment.
A close read of Deepgram's Enhanced STT tier and Future AGI's Agent Command Center gateway, including the documentation gap between the two.
Published July 3, 2026
Reference spec sheet for Deepgram Enhanced, a 2022 speech-to-text tier, including its Future AGI Agent Command Center gateway integration.
Published July 3, 2026
Reference spec for Deepgram Flux, a streaming conversational speech recognition model with built-in turn detection for voice agents.

Published July 3, 2026
What Deepgram Flux actually is: a turn-aware streaming STT model for voice agents, its event API, pricing, benchmarks, and where it falls short.
Published July 3, 2026
Reference profile of Deepgram Nova-3, a proprietary speech-to-text model family for batch and streaming transcription, released February 12, 2025.

Published July 3, 2026
A practitioner's breakdown of Deepgram Nova-3: WER claims, sub-300 ms streaming latency, pricing, languages, deployment options, and where it falls short.

Published July 3, 2026
What Scribe v2 actually changed from v1: features, pricing, benchmark results, API limits, and the architecture details ElevenLabs still won't publish.
Published July 3, 2026
Reference profile of ElevenLabs Scribe v2, a batch speech-to-text model released January 9, 2026: features, benchmarks, pricing, limits, and sources.