5 ms·
the difference is, its not even remotely easy to snap a photo even halfway decent stock photo of an office space with random people, especially with so many peo
by sithlord 4y ago
the difference is, its not even remotely easy to snap a photo even halfway decent stock photo of an office space with random people, especially with so many people not working from home.
it's pretty easy to tell some generative software to "Female giving a presentation in a well lit office with windows open" and get something pretty decent.
This goes for anything that needs props that a normal person doesn't have.
- krisoft 4y ago> it's pretty easy to tell some generative software to "Female giving a presentation in a well lit office with windows open" and get something pretty decent. Is it? I tried that exact prompt with dall-e and got 4 images with nightmare fuel faces: https://imgur.com/a/yyt5Tax https://imgur.com/a/yyt5Tax
- orbital-decay 4y ago>it's pretty easy to tell some generative software to "Female giving a presentation in a well lit office with windows open" and get something pretty decent. As someone who actually tried it, with current models it's possible but neither easy nor straightforward. Getting decent and coherent images is still hard and requires finetuning, so there is certain value in this. And for very specific and series-consistent images I'd still rather set up a photo session if I had a studio. It will change for stock photos, but probably not very soon.
- apetresc 4y ago> it's pretty easy to tell some generative software to "Female giving a presentation in a well lit office with windows open" and get something pretty decent. Tell me you haven't tried Stable Diffusion without telling me you haven't tried Stable Diffusion.
- Jerrrry 4y ago"tell me ___ without telling me" is reddit-style trash trope conversation, please keep it over on reddit, or facebook, or anywhere else, please.
- rjh29 4y agoYou can fix the faces with inpainting. Or Automatic1111's repo has several face restoration tools built in.
- kazinator 4y agoEvery step like that, like image-to-image, adds to how much time you put into it, which is relevant to the buy versus generate yourself question.
- rjh29 4y agoThere is some skill involved but once you know how to use the tools it does not take long to generate a good image and upscale it. If a good stock photo exists I think I'd prefer that, but the real power in AI is to avoid hiring someone to create something original. I agree that anatomy and hands are a problem. Give it 6-12 months and they won't be. That's how fast things are moving at the moment!
- kazinator 4y agoI think what is needed is specialized models. Say that you're in a business in which you just need pictures of residential interior designs and nothing else. A model which is just trained on billions of images of interiors is going to work better than one trained on billions of images of everything.
- rjh29 4y agoAgreed, but strictly speaking the model would probably still be based on a general model such as Stable Diffusion, then "fine-tuned" to render only interior designs. For example the NovelAI anime image generator was based on SD then fine tuned on anime art, and it works incredibly well - it can generate anime art of basically anything you ask, because the model has the wide knowledge of the Stable Diffusion model but is fine tuned to the anime art style.
- kazinator 4y agoIf only this worked: "female with two hands that have five fingers, and normal looking face, giving a presentation in a well-lit office with correct shadows, and a clock on a wall that is perfectly round with legible roman numerals correctly placed around the dial ..."