4 ms·
I mean.. there’s a reason that others haven’t really done this as far as I know and that he only used a tiny sample for a couple seconds: the output from AI voi
by thatswrong0 4y ago
I mean.. there’s a reason that others haven’t really done this as far as I know and that he only used a tiny sample for a couple seconds: the output from AI voice generation generally has a ton of artifacts and isn’t something you’d want to play on a festival sized sound system for more than a few seconds.
So he basically did the only acceptable thing I’d consider doing with this and.. it’s just not that interesting?
- prepend 4y agoIt’s certainly your prerogative to not find it interesting and your feelings are valid. However, I think it’s interesting and can’t wait to see more djs and even new artists using this technique to create some neat art. And of course, it looks like there are tens of thousands of fans in that stadium who seemed kind of interesting. I don’t think statistical sanity is a good measure for value, especially art, but I don’t think it’s noteworthy that you don’t find it interesting.
- thatswrong0 4y agoI understand that. But he's making a big deal out of using literally the two easiest tools out there to generate... a tiny awkward two second soundbite in a set. I just timed myself and it took me less than two minutes to do the same thing. Might take longer if you need to sign up first. And this isn't to say that the time it takes to do something is important - the idea of course matters too. But the idea itself isn't original. And the result is just a TINY soundbite. So it's just straight up not interesting from a production nor listening perspective, especially given the fact there are people out there doing much more creative, involved, and interesting things in the realm of AI assisted production. For example: https://twitter.com/thisiscyclops/status/1619085469198721038?s=20 https://twitter.com/thisiscyclops/status/1619085469198721038... If David Guetta wasn't a big DJ, no one would care about this.
- Miraste 4y agoI don't think this is true any more. At least, I've heard generated audio from the latest models that runs for well over a minute without obvious artifacts.
- thatswrong0 4y agoI don’t know what else I expected to hear since progress in these spaces is moving crazy quickly. What models?
- Miraste 4y agoElevenlabs, by some ex-googlers: https://beta.elevenlabs.io/ https://beta.elevenlabs.io/ If anything, the results I've heard from this are better than the demos. And VALL-E by Microsoft, which isn't quite as good, but notable because it can clone real voices with only three seconds of training data: https://valle-demo.github.io/ https://valle-demo.github.io/