4 ms·
Yeah, it’s basically stable diffusion but for music
by kweingar 3y ago
Yeah, it’s basically stable diffusion but for music
- drummojg 3y agoI used "rollicking" in one description and it was exactly what it sounds like to your unconscious when you are far too drunk at a country bar
- famouswaffles 3y agoNot Stable Diffusion. MusicLM is a neural codec language model. Basically if GPT was predicting the next audio token while conditioned on text.
- ShamelessC 3y agoSo like DALLE-1? FWIW, the various diffusion models (tend to) use cross attention to an attention-based approach.