3 ms·
(Disclaimer: we work on voice synthesis/style transfer and cloning) Depends on how much fidelity you want and how much lag you are willing to accept. Our curre
by narrationbox 6y ago
(Disclaimer: we work on voice synthesis/style transfer and cloning)
Depends on how much fidelity you want and how much lag you are willing to accept. Our current voice style transfer state of the art is sufficiently capable though the results may still need anywhere from 6 months to two years of development to be considered "production ready" i.e think poor quality, noise, and artefacts in output audio with existing tech. Pasini. has a pretty good blog post and paper on this:
https://towardsdatascience.com/voice-translation-and-audio-style-transfer-with-gans-b63d58f61854 https://towardsdatascience.com/voice-translation-and-audio-s...