4 ms·
Man, based on how often Alexa misunderstands me or tries to upsell me on some feature after I get what I want from it, mine is going to sound really pissed off.
by CSSer 4y ago
Man, based on how often Alexa misunderstands me or tries to upsell me on some feature after I get what I want from it, mine is going to sound really pissed off.
I also can’t imagine speaking to my Grandma/pa the way I speak to Alexa either, so that’s food for thought.
Jests aside, this is already possible with ML and a large enough data set. There’s nothing state of the art to see here, right? Just implementing existing tech at Enterprise scale/telling the masses?
For examples, see 15.ai or https://github.com/CorentinJ/Real-Time-Voice-Cloning https://github.com/CorentinJ/Real-Time-Voice-Cloning. There’s another commercialized service similar to the latter example here I saw recently but I can’t recall the name. I wish the article had some more technical details.