Home / Companies / AssemblyAI / Blog / Post Details
Content Deep Dive

Introducing the Dictation API: the first API built for dictation

Blog post from AssemblyAI

Post Details
Company
Date Published
Author
-
Word Count
2,049
Company Posts That Month
12
Language
English
Hacker News Points
-
Post removed?
No
Summary

AssemblyAI has launched its Dictation API, a speech-to-text service designed to turn spoken utterances into ready-to-send text rather than verbatim transcripts by removing filler words, resolving self-corrections, preserving tone, and optionally formatting output for contexts such as task lists, messages, or clinical notes. Built on the company’s Universal-3.5 Pro model, the API supports 19 languages, costs $0.62 per hour of audio, returns both the original transcript and rewritten text, and allows developers to improve recognition of specialized vocabulary through contextual prompts and key-term spelling biases. AssemblyAI reports typical short-request response times below one second and says audio can be uploaded in chunks while a user is speaking; if output processing times out, the service returns the verbatim transcript instead. The company cites a 3.87% normalized word error rate on short-form English benchmark audio, offers a Python SDK and HTTP integration, and has also released Blurt, a free open-source macOS dictation application that demonstrates the API in use.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Real-time 13 649 155 80 -85%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.