3 ms·
Went looking for audio samples, here's some from one of the researchers: http://www.zhizheng.org/demo/is15_mte/demo.html http://www.zhizheng.org/demo/is15_mte/
by julespitt 11y ago
Went looking for audio samples, here's some from one of the researchers:
http://www.zhizheng.org/demo/is15_mte/demo.html http://www.zhizheng.org/demo/is15_mte/demo.html
http://www.zhizheng.org/demo/dnn_tts/demo.html http://www.zhizheng.org/demo/dnn_tts/demo.html
- raverbashing 11y agoAlso here http://homepages.inf.ed.ac.uk/zwu2/demo/icassp16/lstm.html http://homepages.inf.ed.ac.uk/zwu2/demo/icassp16/lstm.html
- sawwit 11y agoI thought this would be about text-to-speech applications, while this seems more like an encoder-decoder problem (make the network learn a pattern and then let it reproduce it). I'm wondering how long it is until we see working TTS based on LSTM RNNs.