4 ms·
If you want to see the most basic form of this pipeline, I have a blog post on "bad speech synthesis" [0]. There are open source versions of WaveNet for TTS [1]
by kastnerkyle 10y ago
If you want to see the most basic form of this pipeline, I have a blog post on "bad speech synthesis" [0]. There are open source versions of WaveNet for TTS [1], but I have not run the code myself or seen the quality of the output. Our code for char2wav is theoretically available [2], but not yet ready for "how-to-guide" level use.
[0] http://kastnerkyle.github.io/posts/bad-speech-synthesis-made-simple/ http://kastnerkyle.github.io/posts/bad-speech-synthesis-made...
[1] https://github.com/buriburisuri/speech-to-text-wavenet https://github.com/buriburisuri/speech-to-text-wavenet
[2] https://github.com/sotelo/parrot https://github.com/sotelo/parrot