Home / Companies / Speechmatics / Blog / Post Details
Content Deep Dive

Boosting sample efficiency through Self-Supervised Learning

Blog post from Speechmatics

Post Details
Company
Date Published
Author
Bethan Thomas
Word Count
1,259
Company Posts That Month
6
Language
English
Hacker News Points
-
Post removed?
No
Summary

We have demonstrated that scaling self-supervised learning significantly improves the sample efficiency of automatic speech recognition (ASR) models. By leveraging large amounts of unlabeled data, these models can learn rich representations of input features and improve performance with fewer samples of labeled data. This approach is particularly effective in low-resource settings where training with high sample efficiency is crucial. Our experiments show that scaling self-supervised learning leads to greater sample efficiency and generally better performance, even when reducing the amount of labeled training data by several orders of magnitude. The results have significant implications for ASR systems, enabling them to achieve excellence with a fraction of the hours of labeled data typically required.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Vector Search 1 638 112 54 -23%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.