4 ms·
There are new, very large public English speech datasets: Mozilla Common Voice, National Speech Corpus, which can be combined with LibriSpeech to train large mo
by bginsburg 6y ago
There are new, very large public English speech datasets: Mozilla Common Voice, National Speech Corpus, which can be combined with LibriSpeech to train large models.
- eindiran 6y agoIf you combine them you get 5ish K hours of speech for English, which is still fairly small compared to what most big players have access to.
- solidasparagus 6y agoAmazon worked with 7k hours of labeled data + 1 million hours of unlabeled data - https://arxiv.org/pdf/1904.01624.pdf https://arxiv.org/pdf/1904.01624.pdf