Blog
Analysis, model profiles and deep dives on speech-to-text.
Analysis, model profiles and deep dives on speech-to-text.
Published July 3, 2026
Reference profile of Amazon Transcribe Medical, the AWS managed API for US English medical speech-to-text, launched in December 2019.

Published July 3, 2026
What Amazon Transcribe Medical offers in 2026: features, pricing vs Google and Nuance, HIPAA posture, research clues, and where the service falls short.
Published July 3, 2026
Reference profile of AssemblyAI Universal-3 Pro: release date, prompting model, language support, pricing, deployment, benchmarks, and disclosed limits.

Published July 3, 2026
What Google Cloud Chirp 3 actually is: release timeline, WER and Elo benchmarks, pricing, known issues, and how it stacks up against Azure and ElevenLabs.
Published July 3, 2026
Reference profile of Google Cloud Chirp 3, a managed speech model family covering multilingual transcription, HD text-to-speech, and instant custom voice.

Published July 3, 2026
Where Deepgram Base fits in 2026: API behavior, variants, latency, concurrency, missing benchmarks, and when to pick Nova-3 or Flux instead.
Published July 3, 2026
Reference profile of Deepgram Base, a legacy speech-to-text model family with task-specific variants, batch and streaming APIs, and self-hosted deployment.

Published July 3, 2026
A close read of Deepgram's Enhanced STT tier and Future AGI's Agent Command Center gateway, including the documentation gap between the two.
Published July 3, 2026
Reference spec sheet for Deepgram Enhanced, a 2022 speech-to-text tier, including its Future AGI Agent Command Center gateway integration.
Published July 3, 2026
Reference spec for Deepgram Flux, a streaming conversational speech recognition model with built-in turn detection for voice agents.

Published July 3, 2026
What Deepgram Flux actually is: a turn-aware streaming STT model for voice agents, its event API, pricing, benchmarks, and where it falls short.
Published July 3, 2026
Reference profile of Deepgram Nova-3, a proprietary speech-to-text model family for batch and streaming transcription, released February 12, 2025.