Introducing the Dictation API: the first API built for dictation
Blog post from AssemblyAI
AssemblyAI has launched its Dictation API, a speech-to-text service designed to turn spoken utterances into ready-to-send text rather than verbatim transcripts by removing filler words, resolving self-corrections, preserving tone, and optionally formatting output for contexts such as task lists, messages, or clinical notes. Built on the company’s Universal-3.5 Pro model, the API supports 19 languages, costs $0.62 per hour of audio, returns both the original transcript and rewritten text, and allows developers to improve recognition of specialized vocabulary through contextual prompts and key-term spelling biases. AssemblyAI reports typical short-request response times below one second and says audio can be uploaded in chunks while a user is speaking; if output processing times out, the service returns the verbatim transcript instead. The company cites a 3.87% normalized word error rate on short-form English benchmark audio, offers a Python SDK and HTTP integration, and has also released Blurt, a free open-source macOS dictation application that demonstrates the API in use.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Real-time | 13 | 649 | 155 | 80 | -85% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.