4 ms·
are there any wav2letter models ready for download ?
by ftreml 7y ago
are there any wav2letter models ready for download ?
- lunixbochs 7y agoFacebook posted some of the models listed in their SOTA paper already. I haven’t tried these yet. https://github.com/facebookresearch/wav2letter/blob/master/recipes/models/sota/2019/README.md https://github.com/facebookresearch/wav2letter/blob/master/r... My models are here: https://talonvoice.com/research/ https://talonvoice.com/research/ I haven’t yet posted the model I’ve been working on most recently. I’m 120 epochs into a large size model trained on all of my datasets. I also have another 1000-1500h of audio I haven’t finished prepping to train on. Here is a web demo. It’s currently running a slightly older checkpoint of my WIP large model, and the deepspeech LM: https://web2letter-west-1.talonvoice.com/ https://web2letter-west-1.talonvoice.com/
- ftreml 7y agoso this is a real cool project. as soon as you finished your training I will be happy to add it as option (or maybe default setup) in the botium speech processing setup (if you want that). Do you have any experience with online decoding in wav2letter ? Is there something like a Websocket API available somewhere ?
- lunixbochs 7y agoHow does accuracy compare to what you’re using now?
- ftreml 7y agothe german model is from the kaldi tuda recipe with WER of 15%. the english is from the tedlium recipe with WER of 7%. room for improvement, but for our original purpose it was sufficient.