3 ms·
Wow, these audio samples are incredible. I'm surprised to hear the model actually outputting natural-sounding breathing between and inside sentences. Most TTS s
by buss 8y ago
Wow, these audio samples are incredible. I'm surprised to hear the model actually outputting natural-sounding breathing between and inside sentences. Most TTS systems explicitly remove things like that, but the addition of breathing makes it sound so much more natural.
The style tokens result in pretty incredible and realistic audio.
- nmstoker 8y agoIf you want to see some more research samples check out this link: https://google.github.io/tacotron/ https://google.github.io/tacotron/ What's especially impressive is how fast they're moving along with new ideas (see the dates) Bear in mind that the WaveNet outputs are likely to be pretty slow to generate (but they do yield remarkable quality!)