iharnoor
8 hours ago
Today we're launching the Dictation API from AssemblyAI, the first ever API built for dictation: your users speak, and it returns a response tailored to your application without any user editing required.
Here's a short clip through our Transcription API versus through the new Dictation API:
- "um so can we uh move the the meeting to thursday i think friday works better actually" - "Can we move the meeting to Friday? That works better."
Traditional transcription models are trained to be verbatim and capture everything you say.
Dictation and Voice Input models need to be aligned for a different task: a cleaned up and properly formatted input to your system or application.
- Voice typing emails and slack messages. - Doctors dictating chart notes. - Voice input to robots and hardware devices. - Coding with your voice.
These are tasks our Dictation API is built for.
We've seen a surge in demand from developers building these types of applications, and the Dictation API now makes it much easier.
It's a simple, low-cost API any developer can use to ship dictation or voice input features and apps within minutes.
- Responses in <200ms - Built on our Universal-3.5 Pro transcription model, ranked #1 for accuracy on independent benchmarks. - Transcribes 30+ languages out of the box and supports language switching. - Include custom formatting instructions on the fly - $0.62 per hour of audio, all in - no token math.
We've also open‑sourced a free Mac dictation app built on the new Dictation API, so you can experience its speed and accuracy firsthand. It's called Blurt.