19 ms·
DreamFusion: Text-to-3D using 2D Diffusion
- MitPitt 4y agoCoincidentally came out the same day as Meta's text-to-video. I wonder if Google deliberately held out the release to make a bigger impact somehow?
- bm-rf 4y agoWould they publish it anonymously? I'd bet they'd want to take credit somehow.
- kmonsen 4y agoSomeone posted the "correct" URL that has names: https://dreamfusion3d.github.io/ https://dreamfusion3d.github.io/
- blondin 4y agonvidia also released GET3D[1] a few days ago. research seems to be heading towards similar goals. [1]: https://github.com/nv-tlabs/GET3D https://github.com/nv-tlabs/GET3D
- astrange 4y agoI think it’s because of ICLR deadlines.
- parasj 4y agoCorrect link with full demo: https://dreamfusion3d.github.io/ https://dreamfusion3d.github.io/
- etaioinshrdlu 4y agoHuh, it's a pretty similar technique to what I outlined a couple days ago: https://news.ycombinator.com/item?id=32965139 https://news.ycombinator.com/item?id=32965139 Although they start with random initialization and a text prompt. It seems to work well. I now see no reason we can't start with image initialization!
- ToJans 4y ago"those who say it cannot be done should not interrupt the people doing it"
- amelius 4y agoThey said it could be done, and even said how...
- efrank3 4y agoThe version that you proposed wouldn't have worked
- bmpoole 4y agodirectly training a NeRF on a single image is a terribly unconstrained problem that would lead to a volume that looks bad when the viewpoint changes. the gist of render + use diffusion model to refine is a great idea though and core to our method! the details of how to use the diffusion model for this refinement was the challenge, but once we figured that out it... just worked :)
- etaioinshrdlu 4y agoVery cool! Can we tweak your algorithm a little to be seeded with a real photo from a single viewpoint?
- poolio 4y agoit's a good idea :)
- 4y ago
- gersh 4y agoIs code available?
- sirianth 4y agoIs there code for any of these models? Or a collab? Ajay Jain's colab doesn't work, but I would love to see a colab for this.
- ajayjain 4y agoHi @sirianth, this is Ajay. Are you talking about the https://ajayj.com/dreamfields https://ajayj.com/dreamfields colab? Feel free to dm me on Twitter (@ajayj_) or email me if you're facing bugs. Dependencies keep shifting, really should have pinned versions originally...
- sirianth 4y agoYes, exactly, I was trying to fix the dependency problems in the colab and I couldn't... I'll hit you up on twitter.
- sirianth 4y agoSilly replying to myself I know, but I had more thoughts. I'm an architect for 3D worlds and I am desperate, lol, for this kind of tool. I use both blender and grasshopper, but I use midjourney to think and prototype all the time. Obvious but it would be astonishing to have something like this for game worlds. I used another version of this to create "a forest emerging from an aircraft carrier" https://www.instagram.com/p/CiRfXKzpnLC/ https://www.instagram.com/p/CiRfXKzpnLC/ but the technique didn't have good resolution yet (high fidelity).
- jaggs 4y agoYou should try the AUTOMATIC1111 version of stable diffusion. It's crazy fast and has great results - https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Install-and-Run-on-NVidia-GPUs https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki...
- sirianth 4y agoI'm mostly only looking for 3DML these days. I want it to hallucinate architecture for games.
- blurbleblurble 4y agoSo does this mean I can use DreamBooth to create plausible NERFs of myself in any scenario? The future is looking weird.
- Filligree 4y agoNah. This is made by Google, so it'll never become useful. You'll have to wait a few months for someone else to replicate it.
- blurbleblurble 4y agoSomeone did implement DreamBooth for stable diffusion so I was imagining, like you say, this (implemented by someone else) in a couple months with DreamBooth + stable diffusion
- RosanaAnaDana 4y agoThis is getting asymptotic.
- layer8 4y agoProgress often happens in waves. There will be a trough again.
- sva_ 4y agoSeems a bit like a tsunami currently. But I wonder how we'll think about it 10 years from now.
- gfodor 4y agoAI might be different - as has been predicted for many years now - due to the compounding effects on intelligence.
- RosanaAnaDana 4y agoI'm not sure this is quite that, if you will. I'm also not sure someone couldn't cleverly figure out a way to use stable diffusion to write code based on a text prompt.
- owenpalmer 4y agoSource?
- golemotron 4y agoAnonymously authored research is very ominous.
- parasj 4y agoThe full author list is on the updated link at: https://dreamfusion3d.github.io/ https://dreamfusion3d.github.io/
- jianshen 4y agoDid we hit some sort of technical inflection point in the last couple of weeks or is this just coincidence that all of these ML papers around high quality procedural generation are just dropping every other day?
- macrolocal 4y agoConference season?
- sirianth 4y agoyay
- deleted 4y ago[deleted]
- layer8 4y agoFrom the abstract: “We introduce a loss based on probability density distillation that enables the use of a 2D diffusion model as a prior for optimization of a parametric image generator. Using this loss in a DeepDream-like procedure, we optimize a randomly-initialized 3D model (a Neural Radiance Field, or NeRF) via gradient descent such that its 2D renderings from random angles achieve a low loss.” This seems like basically plugging a couple of techniques together that already existed, allowing to turn 2D text-to-image into 3D text-to-image.
- blurbleblurble 4y agoTime and time again these ML techniques are proving to be wildly modular and pluggable. Maybe sooner or later someone will build a framework for end to end text-to-effective-ML-architecture that will just plug different things together and optimize them.
- lbotos 4y agoI think this is what huggingface (github for machine learning) is trying with diffusers lib: https://huggingface.co/docs/diffusers/index https://huggingface.co/docs/diffusers/index They have others as well.
- parasj 4y ago@dang The link should be updated to https://dreamfusion3d.github.io https://dreamfusion3d.github.io
- joewhatkins 4y agoThis is crazy good - most prior text-to-3d models produced weird amorphous blobs that would kind of look like the prompt from some angles, but had no actual spatial consistency. Blown away by how quickly this stuff is advancing, even as someone who's relatively cynical about AI art.
- drKarl 4y agoAmazing! How long then until we get photorealistic AI generated 3D VR games and experiences in the metaverse?
- drKarl 4y agoWhy the downvote? I wasn't being sarcastic, it was a honest question, I'm really impressed how far this technology has come since GPT-3 2 years ago to DALl-E and Stable Diffusion ro Meta's text to video to this...
- inerte 4y agoI was wondering the same thing in the other thread about Text to Video. Someone asked about 3D Blender models, which made me think about animating blender models. Bang, now on this thread we see animated images… it does feel like we can get to asking for a 3D environment, put on a VR glass and experience it. And with outpainting, that we can even change it in real time. It’s totally sci-fi, and at the same time seems to be possible? I am amazed how even image generation evolved over the last year, but that’s just me daydreaming.
- macrolime 4y agoOr in-painting with AR glasses. Change things in the real world just by looking at it (with eye tracking) and say what you want it changed into.
- aliswe 4y agoI'm guessing the post could be interpreted as "normie" and not HN curiosity ;)
- rjmunro 4y agoMaybe because you said "Metaverse" (and to some extent "VR") making it sound like sci-fi nonsense. You could have just said: How long then until we get photorealistic AI generated 3D games and experiences?
- macrolime 4y agoThis sounds like something that could be made to work with stable diffusion if someone just implements the code based on the paper.
- coolspot 4y agoGive it a week or two…
- naillo 4y agoIt's funny that the authors are 'anonymous' but they have access to Imagen so obviously it's by Google.
- deleted 4y ago[deleted]
- parasj 4y agoThe full author list is on the updated link at: https://dreamfusion3d.github.io/ https://dreamfusion3d.github.io/
- joewhatkins 4y agoThis is par for the course - there have been other instances where an 'anonymous' paper mentioned training on a cluster of TPUs that weren't publicly available yet - dead giveaway it was Google.
- VikingCoder 4y agoDead giveaway... Dead giveaway...
- googlryas 4y agoLots of reasons to stay anonymous besides for hiding what org is behind the paper. Maybe they don't want to be kidnapped by the North Koreans and forced to produce new paens with "lost footage" to Kim Il-Sung.
- cuuupid 4y agoA large portion of the ML community (rightly) discredits Google papers because: - they rarely provide the data or code used so it's basically "i swear it works bro" research - what they achieve is usually through having the most pristine dataset on the planet and is often unusable by other researchers - other times they publish papers that are basically "we slightly modified this excellent open source paper, slapped an internal name on it and trained it on our proprietary dataset" - sometimes they achieve remarkably little but their papers still get a shiny spot because they're a big name and sponsor all the conferences - they've also been caught trying to patent/copyright ML techniques; disregarding that this is the same as privatizing math, these are often techniques they plainly didn't come up with Also ever since OpenAI did their "we have to go closed-source for-profit to save humanity" PR campaign, every company that releases models that can achieve a large amount in NLP/CV gets dragged by the media and equated to Skynet.
- O__________O 4y agoUnclear to me what is going on, but there’s another URL that lists the authors names. Given it’s possible this change was done for reason, not linking to it, but strikes me as odd it’s still up. Anyone know what’s going on without causing problems for the authors?
- parasj 4y agoThis link was from OpenReview which must be anonymous (double blind). The full author list is on the updated link at: https://dreamfusion3d.github.io https://dreamfusion3d.github.io
- O__________O 4y agoAware of the link, though you have not provided any clarification for why there are two links; strikes me as odd if authors are trying to post it anonymously that simple Google finds authors names.
- rirarobo 4y agoI believe ICLR guidelines require the authors to submit papers and any supplementary materials (including links to webpages, videos, etc) without identifying information, but authors are not barred from public announcements on other forums. IIUC, the idea behind this policy was originally to accommodate author freedom to engage in common practices such as simultaneous submission to arxiv (which identifies the authors). To respect the double blind spirit of review, reviewers are asked not to actively search the web in attempt identify the authors. In the past, when social media promotion was less common, it was reasonably likely that reviewers would follow this guidance and would not have seen the arxiv submission, preserving the double blind nature of review in most cases. However, the use of social media in academia has radically changed in recent years, as more researchers use social media to keep up with the latest advancements, so promotion of papers in submission on platforms like Twitter can offer significant advantages to authors. So, authors today often submit anonymously following the conference guidelines, but simultaneously post publicly elsewhere, walking a fine line as not to overstep the conference policies. This appears to be the case for this submission. Note, recently, some conferences, such as CVPR, have started to institute new policies forbidding social media promotion until acceptance, as they adapt to the changing landscape of social media promotion. If this were a CVPR submission, the authors would not be allowed to tweet publicly about their work yet, nor have the version of the webpage with their names visible.
- ml_basics 4y agoAwesome! I wonder how long it will be until there is an open source implementation compatible with Stable Diffusion
- achr2 4y agoThe thing that frightens me is that we are rapidly reaching broad humanity disrupting ML technologies without any of the social or societal frameworks to cope with it.
- narrator 4y agoThere were bigger disruptions in the past. The telegraph, railroads, explosives. "The Devils" by Dostoevsky is a great fictional account of what all these technological disruptions do to the fragile social order in the late 19th century Russia countryside. All of a sudden all these foreign people, ideas , technology and commerce start streaming in to these once isolated communities.
- throwaway675309 4y agoI'm usually not a fan of this general hand wringing / fear mongering around ML that a lot of people with too much time and not enough STEM background constantly bring up. Stable diffusion has been made available to the public for quite a while now and if anything has disproved a lot of the ungrounded nonsense that made companies like OpenAI censor their generative models.
- deltasevennine 4y ago>I'm usually not a fan of this general hand wringing / fear mongering around ML that a lot of people with too much time and not enough STEM background constantly bring up. What's with the STEM reference? Are you implying that STEM is related to intelligence and that people without a STEM background are not intelligent? It is well known among academics (AKA STEM MAJORS) that human society is a chaotic system and that ML can change society for the better or for the worse... the outcome is basically unknown... It is therefore the intelligent choice to consider the negative consequences of this technology. To not consider the other side indicates a lack of something.
- lolspace 4y agoThat's exactly what he's saying. How many citations does the top cited paper from the LessWrong community have?
- jonas21 4y agoCan someone explain what's going on in this example from the gallery? The prompt is "a humanoid robot using a rolling pin to roll out dough": https://dreamfusion-cdn.ajayj.com/gallery_sept28/crf20/a_DSLR_photo_of_a_humanoid_robot_using_a_rolling_pin_to_roll_out_dough.mp4 https://dreamfusion-cdn.ajayj.com/gallery_sept28/crf20/a_DSL... But if you look closely, the pin looks like it's actually rolling across the dough as the camera orbits.
- WithinReason 4y agoThe rolling pin is above the table but the shading is wrong because they don't render shadows.
- jnbrrn 4y agoI think what's happening here is that the flat-looking table is actually raised up in the center, in the shape of something like a smooth pyramid. There's dough painted on both sides of the rolling pin, but because of the curvature of the "table" you only see each side's dough when the camera is on that side of the pyramid.
- modeless 4y agoThe most incredible thing here is that this demonstrates a level of 3D understanding that I didn't believe existed in 2D image models yet. All of the 3D information in the output was inferred from the training set, which is exclusively uncurated and unsorted 2D still images. No 3D models, no camera parameters, no depth maps. No information about picture content other than a text label (scraped from the web and often incorrect!). From a pile of random undifferentiated images the model has learned the detailed 3D structure and plausible poses and variants of thousands (millions?) of everyday objects. And all we needed to get that 3D information out of the model was the right sampling procedure.
- adamredwoods 4y agoSo I wonder if unusual angles that normally do not get photographed will be distorted? For example, underneath a table looking up.
- jacobr1 4y agoThey reapply noise to the potentially distorted image and then predict the de-noised version like the originally rendered first frame. So the image is at least internally consistent for the frame (to the extend the the system generates consistency whatsoever). The example with a squirrel wearing a hoodie demonstrates an interesting edge case, the "front" of the squirrel (with hoodie over the head) show a normal hooded face as expected, but when you rotate to the "back" you get another face where the hoodie is low over the eyes. Each looks fine in isolation, but in aggregate it seems like we have a two-faced squirrel.
- yarg 4y agoIt'll be delusions and guesses, rather than distortions. It'll just make up some colours and geometries that don't contradict anything it already knows from the defined perspectives. Or leave it empty.
- poolio 4y agoYes, this is often a problem. We use view-dependent prompts (e.g. "cat wearing sunglasses, back view") but the pretrained 2D model often does not do a good job of interpreting non-canonical views and will put sunglasses on the back of the cats head (as well as the front).
- samuell 4y agoGives a new perspective on a classic verse: "For he spoke, and it came to be; he commanded, and it stood firm." Psalm 33:9, NIV :)
- dang 4y agoUrl changed from https://dreamfusionpaper.github.io/ https://dreamfusionpaper.github.io/ to the page that names the authors.
- VikingCoder 4y agoWe're quickly approaching HNFusion: Text-to-HN-Article-That-Implements-That-Idea ...
- yarg 4y agoCool. The samples are lacking definition, but they're otherwise spatially stable across perspectives. That's something that's been struggled with for years.
- arisAlexis 4y agoIs it a light version of script when the AGI comes fast
- coolca 4y agoThis is like magic to me. The pace at which we are getting these tool amazes me.
- edgartaor 4y agoI don't see a person in the gallery. It's capable of generate a 3D model of me with only a photo?
- WheelsAtLarge 4y agoI'm not even going to pretend that I have a clue on how this is done. But I'm wondering if the output can be turned into 3d objects that can be used in any of the 3D modeling software? It would be a game changer in terms of real world product development in both of speed and ease.
- hwers 4y agoWell they even have a “downlod model” so yep you definitely can. I wouldn’t think of this as an amazing panacea though, since once everyone has access to it that suddenly means whatever reason making assets like this was valuable before, will now be dirt cheap for all and thus actually net negative for people in that industry. Just saying and warning, not to be a bummer
- WheelsAtLarge 4y agoThx, I went back and saw that I missed this the first time: "Mesh exports Our generated NeRF models can be exported to meshes using the marching cubes algorithm for easy integration into 3D renderers or modeling software." Like they say, "This is the start of something big."
- cpitman 4y agoFWIW, there's still a pretty big gap between a single static mesh and something that is a usable asset, say in a game. Maybe this could provide a shortcut for a modeler to get started, but it still is going to take a lot of skill from that point.
- birracerveza 4y agoIn the make-a-video I said that things are getting more and more impressive by the day. I was wrong, because that was a couple hours ago. They're getting more and more impressive by the HOUR. I'm curious where this will end up in a year. Will it plateau? If so, when?
- keepquestioning 4y agoOh my god, we are done for
- samstave 4y agoAs someone who went to college for 3D animation in +*1997*+ AND DESIGNED the datacenter for luca' presidio complex.. where-by learning that Pixar was developed by steve jobs when lucas didnt think there was a future for computer animation... and so steve bought the death star from lucas... That became pixar... AI is going to fucking kill it - what will happen in the next decade will be ANYONE uploading a script to an AI to make a full length movie... AND their will be editing tools as well that are AI driven... Like mentioned by William Gibson *The future is here, its just not evenly distributed yet*
- mlajtos 4y ago> AI is going to fucking kill it - what will happen in the next decade will be ANYONE uploading a script to an AI to make a full length movie... Nah. These techniques will definitelly lower the barrier for making stuff (not just movies), but that has been case will all transformative technologies. Before computers, if you wanted to shoot and edit a movie, it was a challenge. Now, you can shoot a movie with your pocket computer and edit it on the same device while shitting. Upload it to video sharing web and billions of people can watch it. This class of technologies will enable creative people to make a lot of stuff, to iterate quickly. But don’t be naïve that everybody will do that. 99% of that will be trash, and that is fine. I like that it will enable individuals to bring their visions into this world without any need for collaboration. And when highly artistic individuals will begin to collaborate using these tools, that will be an awesome inflection point for art as we know it.
- deltasevennine 4y agoIt's going to be worse then that. I'll just write a summary and an AI will generate the full script. Then the movie will be generated from the script. The full source code and assets for a video game too. All the primitive components for this future seem to be at an early stage of inception. We can't say for sure whether they will mature to the point where they can replace us, but the trajectory is certainly pointing in that direction.
- dbspin 4y agoI think this will come, but it won't be competitive with filmmaking in a decade, not without real superhuman AGI. What's much more likely is efforts to reproduce performances will result in deep uncanny valley stuff - and correcting this last few percent for accuracy / weirdness will take a long time. Photorealistic video rendering is inevitable, voice duplication is already here (Tortoise should be much more well known but our culture is fixated on the visual side of this tech - https://colab.research.google.com/drive/1wVVqUPqwiDBUVeWWOUNglpGhU3hg_cbR https://colab.research.google.com/drive/1wVVqUPqwiDBUVeWWOUN...). But generatively creating a performance which accurately interprets the emotional context of a scene, in interrelationship with other characters doing the same? I don't see even a path towards that without AGI. At best you could crib specific performance elements and map them onto new models. My intuition is that to get all the way to generative movie you need AGI (or Zimbos so convincing we can't tell the difference). So what we will get in coming years / decades - the timeline is anyones guess - is movies acted out to camera in a rehearsal space, and that combined with a(n AI augmented) script to generatively create a film / character driven interactive game).
- bmpoole 4y agohi folks, ben p from the dreamfusion paper here. happy to answer qs for the next ~hour!
- suyash 4y agoIs there code or notebook for this paper?
- bmpoole 4y agonot yet, but see the appendix of the paper for pseudocode. the core update step from the diffusion model that powers dreamfusion is surprisingly simple and easy to implement.
- thinkingemote 4y agoCurious about how long it took: brain storming, research, hypothesis, work, iteration, bug fixing, writing etc I'm curious about the process you and the team uses to do this. Additionally there's the meme that these things are appearing every hour now, so it could be good for some perspective like "well actually it took n weeks"
- poolio 4y agoGreat question! Our team has been working on text-to-3d for ~1.5 years starting with https://ajayj.com/dreamfields https://ajayj.com/dreamfields. We had hoped that we could swap the contrastive CLIP model in Dream Fields for the generative Imagen model and crank out an easy paper in a few weeks. But what was supposed to be an easy win turned into months of frustration. Nothing we tried worked any better than Dream Fields. After a long detour trying MCMC, we stumbled across the score distillation loss that powers DreamFusion. Going from an initial sign of life to the results you see today still took months of hard work. Research progress is unpredictable and these advances are not inevitable. We have the privilege to work in an environment full of amazing colleagues and powerful models, but at the end of the day it took a persistent team and a bit of luck.
- bertdb 4y ago
- visarga 4y agoFuturists have been predicting when we'll have stable fusion for decades, but now we suddenly got stable diffusion working. That's good too, not what we wanted, but good. We're gonna need stable fusion or other renewables to run stable diffusion though. /s
- MrLeap 4y agoFun that they had an octopus playing a piano. I made the same thing the old fashioned way. Mine can actually play though. https://twitter.com/LeapJosh/status/1423052486760411136 https://twitter.com/LeapJosh/status/1423052486760411136 :P
- xixixao 4y agoText to 3D animation is the obvious next step though.
- deltasevennine 4y agoWhat does this mean for our understanding of intelligence? It trivializes it, in my opinion. When asked the question of is lambda/GPT-3 and/or DreamFusion and it's derivatives an aspect of sentience? there's always a bunch of people who are repeating the same cliche negative line, of "no, it's only attempting to statistically mimic sentience." I agree with the reasoning. But have we considered the other side of the story? That yes, the mimicry is All sentience actually is. Nothing more.
- mattnewport 4y agoAI is getting quite good at a lot of things humans consider fairly difficult (like this example) but has made less progress at things humans consider fairly easy (e.g. maintaining consistency across paragraphs of text for GPT3, navigating 3D space, learning from small numbers of examples). That suggests that there is still a gap between what current AI approaches are doing and what the human brain is doing that doesn't just come down to throwing ever more data at the problem.
- deltasevennine 4y agoYou just picked an arbitrary gap though. And it seems like a small gap that's crossable. For example just 2 weeks ago there were two other gaps that you could've used in your example. You could've said 3D interpretation of images wasn't possible and the creation of animated movies wasn't possible and you could've said these few things suggest that there's a gap in what the human brain is doing and what the AI is doing and just throwing more data at the problem doesn't fix it. Those two examples would be irrelevant today as both of those gaps have Effectively been crossed. See what I'm saying here. There's two ways of looking at it even from your perspective... Either that gap is so large that the human brain is completely different. Or the gap is small, trivial and will be crossed very very soon.
- mattnewport 4y agoI'm saying something different, that the most impressive examples of AI breakthroughs are doing things that humans find hard / are bad at. Meanwhile there are many things that people find easy / do without thinking / can be done by dogs or very young children that AI struggles with. It suggests to me that what most current approaches are doing is something fairly different from what human / animal intelligence is doing in important ways. That means we will likely continue to see AI do increasingly amazing things while at the same time struggling to perform a lot of tasks that are quite basic for humans. It is the fact that AI is proving to be a better artist than most humans while not being able to do many things that are simple for a 4 year old that suggests strongly to me that some of the fundamental mechanisms are fairly different still, or current AI approaches are missing some key insights. I could be wrong. I'd bet money that I'm right if there was an easy way to do it though.
- Dirgatara 4y ago
- LarsDu88 4y agoAs someone who dabbles in 3d modeling, this is going to be an incredible resource for creating static 3d objects. Someone ought to come up with a way to convert to mesh better than the Marching Cubes algorithm I've seen applied to most NERFs. The models still lack coherent topology and would probably be janky if fully rigged.
- boppo1 4y agoI swear I recently saw something related to generating clean topology procedurally. Wish I could remember where.
- poolio 4y agoWith smooth enough geometry converting NeRFs to meshes with marching cubes works pretty well. Would you say the topology of meshes on our website are still too incoherent for rigging?
- hjaarnio 4y agoIn case of a communication gap; The word 'topology' has a more domain specific meaning in animation and rigging compared to the mathematical one. It's used to mean that the placement of the lower level components - vertices, edges and faces - are well aligned to the higher level structure of the object, and make up a well defined 2D grid that flows along the models surface. In particular you'd want edges going along/perpendicular to mesh structures such as limbs, and around facial features and other details in a logical manner. Otherwise when applying deformations as part of an animation the model will not have the detail in the right places to still look good, e.g. if there is no edges perpendicular to a joint in a limb, the bent version of the limb cannot have a clear smooth line along the joint, and the edges and faces become janky. Under this definition, marching cubes cannot produce good 2D topology, as the mesh features are all aligned to the cardinal grid instead of the features of the object represented.
- poolio 4y agoAha thank you, this is helpful. Agreed there is much research needed to get this working but hopefully not too far off: https://twitter.com/kkpatain/status/1575758085821706240 https://twitter.com/kkpatain/status/1575758085821706240
- EZ-Cheeze 4y agobtw guys stable diffusion img2img consistently applied frame-by-frame will get us some insane CGI for movies yo "transform this into this realistically" ILM's holy grail
- tonis2 4y agoIs there an API for using it myself ?
- kennyloginz 4y agoPretty neat, wish I could try it out ( maybe I missed a link). Obviously has interesting / novel uses, but kind of reminds me of the previous discussion of upscaling audio recordings to the “soundstage” format. I doubt most 2d images want to be 3d ;)
- dusted 4y agothat is so amazing! Next up it puts a skeleton in them and animate :o
- spaceman_2020 4y agoThese are getting too good, too fast. I'm excited and scared. The world is going to look very different in 10 years!
- airbreather 4y agobut seems i can only generate models from predetermined inputs, when can i submit my inputs to create a video?
- deleted 4y ago[deleted]