January 2021 Summaries
1 posts from AssemblyAI
Filter
Month:
Year:
Post Summaries
Back to Blog
AssemblyAI's research and development efforts focus on improving the accuracy of their Speech-to-Text API. They are exploring new architectures for end-to-end speech recognition, such as Listen Attend and Spell (LAS) and Recurrent Neural Network Transducers (RNNT). These models have shown production level accuracy matching or surpassing that of conventional hybrid DNN-HMM systems. The LAS model is perceived to have better accuracy than RNNT, but RNNT models are seen as having more desirable features for production use. Combining LAS and RNNT can achieve better accuracy and feature parity when compared to hybrid RNN-HMM models.
Jan 01, 2021
3,372 words in the original blog post.