3 ms·
Only for images. People want to generate videos next and those models will be likely GPT-sized.
by bitL 4y ago
Only for images. People want to generate videos next and those models will be likely GPT-sized.
- Metus 4y agoThere is a video model making the rounds on /r/stablediffusion and it is just a tiny bit larger than Stable Diffusion.
- isoprophlex 4y agoYou're not kidding! it's far from perfect, but pretty funny still... https://www.reddit.com/r/StableDiffusion/comments/126xsxu/nightmare_continues_octopus_dinner https://www.reddit.com/r/StableDiffusion/comments/126xsxu/ni... Too bad SD learned the Shutterstock watermark so well, lol
- bitL 4y agoIt's cool though not very stable in details over temporal axis.
- Metus 4y agoOf course the quality is horrible relative to a proper video, it just illustrates that txt2vid might not need 100B+ parameters.