5 ms·
Goodness me this is the stuff of nightmares! I asked it for "A dog catching a treat in slow motion, it's chops flapping around comically" https://imgur.com/a/
by pmx 2y ago
Goodness me this is the stuff of nightmares!
I asked it for "A dog catching a treat in slow motion, it's chops flapping around comically"
https://imgur.com/a/dog-catching-treat-slow-motion-its-chops-flapping-around-comically-6VAyL3O https://imgur.com/a/dog-catching-treat-slow-motion-its-chops...
- usful 2y agoSame nightmare fuel for me https://imgur.com/a/warriors-who-are-cows-fighting-with-edo-period-japanese-weapons-on-open-meadow-camera-tracking-alongside-them-smooth-side-angle-motion-WXgXAR2 https://imgur.com/a/warriors-who-are-cows-fighting-with-edo-...
- rpastuszak 2y agoI'm reading the Southern Reach Trilogy at the moment, and this type of imagery fits there really well.
- seydor 2y agoNightmares are also dreams
- csomar 2y agoI can see him catching the treat (middle of the video) in super fast motion.
- ilaksh 2y agoI wonder if this is doing basically the same thing as SORA/Kling, but with less compute/model size. It kind of reminds me of OpenAI's examples from their technical report: https://openai.com/index/video-generation-models-as-world-simulators/ https://openai.com/index/video-generation-models-as-world-si... Somewhere between "base compute" and "4x compute". So maybe you "just" need to know how to create a certain type of diffusion transformer model and then train on a ton of videos, but with an adequate amount of compute. Which is probably a LOT for training and inference to get more realistic results.
- deleted 2y ago[deleted]