4 ms·
Cool project, seems like your model have similar WER as mine (4th reference in readme). Do you plan to do any pre-training on the encoder part in the future? Ma
by blackcat201 6y ago
Cool project, seems like your model have similar WER as mine (4th reference in readme). Do you plan to do any pre-training on the encoder part in the future? Maybe something like this[1]
[1] https://ai.facebook.com/blog/wav2vec-20-learning-the-structure-of-speech-from-raw-audio/ https://ai.facebook.com/blog/wav2vec-20-learning-the-structu...
- iceychris 6y agoHey blackcat! Your project [0] helped me a lot! Pre-training the encoder sounds great, I'll maybe add it in the future. [0] https://github.com/theblackcat102/Online-Speech-Recognition https://github.com/theblackcat102/Online-Speech-Recognition
- bravura 6y agoHey black cat, I have some work in preprint for a NeuroIPS workshop, demonstrating negative results of different audio distances on pitch tasks. There is one particular w2v result I'd like your feedback on. Do you mind emailing me? Lastname at gmail dot com (see my profile for my name)