11 ms·
I recreated famous album covers with DALL-E
- kaffeeringe 4y agoI wonder, how much energy is being burned for these kinds of experiments.
- lalopalota 4y agoprobably less than the amount of energy being burned by people browsing hn.
- teddyh 4y ago> The question is, when the music blows up and the artwork becomes a signature, like the Rolling Stones' Tongue & Lips, who will own the copyright? That’s what trademarks are for.
- sieabah 4y agoIf they had bothered to read the license agreement they would know whoever generates the art owns the copyright. Since it's a pay-for action the copyright is owned by the payee. So really they generated these and never bothered to do the research of their own question.
- teddyh 4y agoI have, like the article author presumably also does, a profund doubt as to whether generated works of this kind can be free of any copyright as long as the tool used is itself created using myriads of copyrighted works (as training data). I certainly do not trust the claims of the tool creators; they have all the incentive to ignore any copyright problems in order to get a tool which is usable. And, as the article states: > But seriously, how creative and original can you be with something that is trained on the works of millions of other creators? > To me, it is unclear whether you can actually call these works your 'own' at all, because there's always someone else's touch on it. > […] users of DALL-E will also never be sure whether they are generating something that is 'theirs' or just a knockoff of someone else's work.
- hedora 4y agoOK, so I give you license to use this URL I just generated to generate your own stuff. It's pay for action (send me a penny if you find anything worthwhile), and the copyright is owned by the payee: https://images.google.com/ https://images.google.com/
- meowkit 4y agoThe URL has not been "generated" in the same sense. You are retrieving an existing string. The images from google are not "generated" in the same sense, they are indexed from google's search algorithm. The generative models, specifically for DALLE here, compute pixel unique images. You might say these models index a subset of an extremely high dimensional space (pixel count * RGB color values) using a query. Traditional search engines build an index from nothing and then use a search query to find the best matches in a more discrete space.
- teddyh 4y agoIf I made a website where you could type text and get images, and I said that your held the copyright to any images you got, could you safely act on that assumption? What if my web site was simply a proxy for Google image search?
- andreyk 4y agoI wonder how long the novelty of DALL-E will persist. HN seems to upvote anything titled "I did X with DALL-E". This is a fun post, but it's not that interesting or surprising. Still worth a look don't get me wrong, but personally didn't learn anything new from it. (eg recreating the famous pink Floyd cover with "Outline of prism on a black background in the middle of scene splits a beam of light coming from the left side into rainbow on the right side" unsurprisingly worked somewhat well).
- rjtavares 4y agoJust like images after text, pretty sure once the novelty of images ends we'll move to animations, music, movies, computer games, and so on.
- Apocryphon 4y agoFor the past few years, I’ve long considered Hollywood failed adaptations to be a non-issue because so long as industrial civilization exists, media companies are going to squeeze what they can out of IPs. So while Game of Thrones might have ended quite poorly, I figure we’re only a few decades away from a different rights-holder giving it another go. Thus, I look forward to the AI-generated versions of famous works that deepfake the original cast into speaking (hopefully) better-written dialogue. Imagine when this technology is widespread, fanfic authors rendering their interpretation of works with the descendants of DALL-E. Everyone gets the dream adaptations and sequels and finales they want.
- voltaireodactyl 4y agoThe trouble with this, for me, is that wouldn’t such a world necessarily mean the end of new actors (not to mention new worlds and stories)? After all, why hire unknowns when you can stack the cast with the A-listers of all time? And if everyone stacks the cast thusly, there are no opportunities for new actors to work with old and be mentored + opportunity to simply works since every bit of work is what adds up to make a journeyperson.
- machinekob 4y agoIs i do smth with DALL-E auto top hacker news post i saw like 20 post like that in past 2 weeks.
- Michelangelo11 4y agoMan, after seeing Stable Diffusion's output, DALL-E's looks just janky. Like watching a propeller plane after seeing a jet. Crazy how fast the tech is moving.
- twostorytower 4y agoDALL-E is capable of very high quality photorealistic images with the right prompts. Here's one I made: https://imgur.com/yAzKkHb https://imgur.com/yAzKkHb “High detail, macro portrait photo, a handsome Australian man with a strong jaw line, blue eyes and brown hair, smiles at the camera, set in an outdoor pub at golden hour, shot using a ZEISS Supreme Prime lens”
- Terretta 4y agoThis particular prompt form, including the lens type, is from this reddit thread: https://reddit.com/r/dalle2/comments/wsi97q/_/ikyjqhh/?context=1 https://reddit.com/r/dalle2/comments/wsi97q/_/ikyjqhh/?conte... > High detail, macro portrait photo, a [physical descriptor, regional identity, etc] man/woman with [eye color] eyes and [hair color] hair, smiles at the camera, set in a [field/dimly lit room/whatever] at golden hour, shot using a ZEISS Supreme Prime lens The suggestions in that thread are quite effective, particularly the notion of reducing synthetic ‘beauty’ for a more human appearance.
- svnpenn 4y ago> shot using a ZEISS Supreme Prime lens that seems too specific. You wouldn't even request that in real life, unless you were trying to be ironically pretentious.
- xwdv 4y agoAlthough AI artists will destroy a lot of jobs, it will also create demand for new jobs for people who specialize in “paint overs” – taking a high concept output created by AI artists and touching it up to perfection. Or perhaps even beyond just a paint over, and into the realm of recreating an entire AI artwork but with a human touch to get details just right. Looking forward to it.
- teddyh 4y agoThe new copyright washing industry is nearly upon us.
- soganess 4y agoThis is the most underrated/prescient comment I've seen on hn. Once prompt engineering become a mature field this is going to be a serious issue. Finally, the crossover of creative writing x cs... for graphic design? I can't wait to watch the lawyers recoil. ::Prepares Popcorn::
- avian 4y ago> taking a high concept output created by AI artists and touching it up to perfection. When you put it like that, it sounds like a nightmare up side down world. It's not AI that's the tool for enhancing human creativity. It's humans that are the AI's tool, cleaning up the edge cases the AI artist can't handle (yet). It destroys creative jobs that give joy to people and creates assembly line jobs for them to slog in.
- krapp 4y agoThe purpose of jobs has never been to give people joy, but to extract value from their labor as quickly and efficiently as possible. Getting any kind of emotional satisfaction from one's work is a privilege which arguably points to an inefficiency in the market, as that energy is wasted which could better be put to productivity. Artists, programmers and everyone else will have to find their joy somewhere other than than selling themselves to a corporation, once AI driven markets optimize away any room for "joy" and the like, and that's going to be one of the few good things about automation. The sooner we break people from the Puritan delusion that work defines a person's meaning and the value of their expression, the sooner we can once again decouple culture from the machinery of capitalism.
- sgt101 4y agoLook, it's trained on these images. It's really great and cool and all - but it's retrieving things that it was trained on. Show me something original it did.
- Engineering-MD 4y agoIt’s hard when it was trained on everything pretty much. That’s the same problem as with GPT3. In my mind it’s still brute forcing a solution but instead of endless computation it’s endless examples
- kgeist 4y agoIs this "bruteforcing" really different from what our brains do? We see thousands, millions of little things (examples) every day. Then we combine what we've seen into something new. Probably the only difference is that DALLE's training was done once while our brains are trained every day for 80+ years.
- Engineering-MD 4y agoI would say yes it is. For humans there is plenty in the world that is completely novel, and we can (and have to) reason from abstractions or first principles. Being exposed to millions of every day things doesn’t give us intrinsic knowledge of writing fiction, algebra, artistic expression etc. instead, we have to apply abstract knowledge and reason. This may well be possible for an AI system, but I haven’t seen GPT3 or DALLE do this. It’s hard to test this on DALLE or GPT3 as the internet is a summation of all known knowledge in effect. True unbounded Original thought is hard, and given they have seen everything before, it’s impossible to know if it’s original or just seen it previously. It would be interesting in a decades time to see how DALLE or GPT3 deals with novel ideas that it was never exposed to.
- l33tman 4y agoNone of the AI generators retrieve things they were trained on, they don't work that way. Everything is original. However our definition of "original" might vary a bit, but so it will vary for any work of art any human artist do as well, as they are also trained on the same images. In the end, a lawsuit and a courtroom might have to decide if by chance someone or some AI creates a picture used commercially that seems similar to someone else's trademark or copyright. Most of the images I've generated using Dalle 2 feels completely original. Just have a look at the reddit r/dalle2 and I'm pretty sure you'll also decide they're "original works".
- cowmix 4y agoAfter getting access to the beta, combined with all these HN posts -- I've determined DallE2 is neat but no where as great as the initial samples made me believe.
- twostorytower 4y agoIt is actually incredibly capable but if you're looking for photorealistic images of people, it needs very specific directions. I learned a lot from this person creating AI portrait photography: https://old.reddit.com/r/dalle2/comments/wsi97q/some_of_my_portraits_of_people_that_dont_exist/ https://old.reddit.com/r/dalle2/comments/wsi97q/some_of_my_p...
- yummybear 4y agoI love seeing people experiment with this technology. You can feel we’re on the cusp of something great - whatever it is, we’re just not quite there yet.
- soneca 4y agoHave anyone given a prompt to Dall-e of designing a company website and included “make it pop!”? Maybe the AI will finally get what designers always complained about annoying clients.
- codetrotter 4y agoPrompt: > Create a website design for a company that sells propane and propane accessories. The name of the company is Strickland Propane, a local propane dealership. Make it pop. Results: * https://i.imgur.com/Jv7NJEN.png https://i.imgur.com/Jv7NJEN.png * https://i.imgur.com/5Uiyg1R.png https://i.imgur.com/5Uiyg1R.png * https://i.imgur.com/LL1DC11.png https://i.imgur.com/LL1DC11.png * https://i.imgur.com/buv5BvS.png https://i.imgur.com/buv5BvS.png So there you have it :p
- codetrotter 4y agoAnother one. Prompt: > Create a website design for ACME Corporation, a company which produces a wide array of products that are dangerous, unreliable or preposterous. Include customer quotes from a dissatisfied Wile E. Coyote prominently on the page. Make it pop. Results: * https://i.imgur.com/WK3QBj9.png https://i.imgur.com/WK3QBj9.png * https://i.imgur.com/Bghgzjt.png https://i.imgur.com/Bghgzjt.png * https://i.imgur.com/XLyYx76.png https://i.imgur.com/XLyYx76.png * https://i.imgur.com/QTSyFTc.png https://i.imgur.com/QTSyFTc.png
- codetrotter 4y agoMore. Prompt: > Website design for Weyland-Yutani Corporation. The Company was founded in 2099 by the merger of Weyland Corp and Yutani Corporation. Weyland-Yutani is primarily a technology supplier, manufacturing synthetics, starships and computers for a wide range of industrial and commercial clients, making them a household name. The website design for The Company is mobile first. Make it pop. Results: * https://i.imgur.com/JyhYK5b.png https://i.imgur.com/JyhYK5b.png * https://i.imgur.com/J5aPXCH.png https://i.imgur.com/J5aPXCH.png * https://i.imgur.com/ksqrW09.png https://i.imgur.com/ksqrW09.png * https://i.imgur.com/uXBcGa5.png https://i.imgur.com/uXBcGa5.png
- system2 4y agoDALL-E still seems very useless. Reminds me of the hype of Cardano.
- bryanrasmussen 4y agoNo Smell the Glove cover, this is a black day for rock and roll!
- deleted 4y ago[deleted]
- powersnail 4y agoDALL-E is still highly probabilistic in its judgement. For instance, in this article, it keeps putting "fire" in the the background on something that is likely to be on fire, rather than lighting up the person. I have a similar experience. In my own experiment, I can't get DALL-E to turn off the street lamp at a bus stop in the darkness. I've tried "no light", "broken street lamp", etc.; no use. Any mention of "street lamp", the scene will include a working street lamp. It's just more probable that a scene with a lamp in the darkness must have that lamp providing light, and this is something that DALL-E will not break out of.
- whirlwin 4y agoI have experienced violent or harmful settings to be avoided by DALL-E. E.g. setting a person on fire. Same with drowning - seems to be impossible/hard to generate
- zaik 4y agoViolent images likely have not been part of the training data for obvious reasons.
- thematrixturtle 4y agoAnd this likely also explains why the OP had a hard time generating pictures of naked babies.
- seesaw 4y agoI gave a prompt about a kid reading the Harry Potter book in the bed. It generated a kid wearing Harry Potter glasses reading a book. Pretty close, but also quite different from what I meant.
- joshschreuder 4y agoI asked for "The Walking Dead directed by x" and got a content violation, I guess because my prompt included "dead".
- deleted 4y ago[deleted]
- wodenokoto 4y agoI'd love to see what it had come up with if simply prompted for "Album cover for Nevermind by Nirvana"
- nprateem 4y agoAn upvote for whoever can give me a prompt to generate an image of someone who's been massaged so much their body has been flattened, as if they were made of dough or jelly or something. I spent ages on this earlier getting nowhere. I'm starting to think DALL-E is better if you don't really know what you want and you're just fishing for ideas.
- _the_special_ 4y ago> an image of someone who's been massaged so much their body has been flattened, as if they were made of dough or jelly do you want a realistic looking one? 3d rendered? what do you have in mind exactly?
- birdyrooster 4y agoalso what is the PCs budget?
- nprateem 4y agoAnything really. I tried cartoons, digital art, watercolours, pixar style. None worked.
- fzfaa 4y agoI didn't know that I needed this.
- doerinrw 4y agoOk you owe me $3! This is a really hard prompt, and only got close-ish with inpainting. Got the base figure with "massaged relaxed flattened person, flat, flat, flat, flat, claymation", then finally got it to add a not-too-terrifying face with "photograph of smiling white woman laying on the ground, promotional photography". Final tweaks to erase some artifacts (it really didn't want to believe the figure on the left was the referenced woman) was "photograph of a wooden floor with a white mat and small plants, overhead shot". DALLE is hard! Curious to see if I can be beat. https://imgur.com/a/tuyGjxp https://imgur.com/a/tuyGjxp
- 4y ago
- google234123 4y agoIs the issue with faces a deliberate choice by the devs?
- mcintyre1994 4y agoIn the case of celebrities yep. It can generate original photorealistic (or whatever style) faces, but they won’t let you generate the faces of real public figures AFAIK.
- randymy 4y agoWorth noting that DALL-E automatically “rejects attempts to create the likeness of any public figures, including celebrities". So, you wouldn't be able to get an image that included the 4 Liverpudlians. It does allow you to create fake faces. Might be fun to try and recreate Miles Davis Tutu, Aladdin Sane, Piano Man.
- deleted 4y ago[deleted]
- cameronh90 4y agoMy experience was that if you name a celebrity (and the request isn't blocked) it quite often generated something that has the same general vibe of the target, while also looking entirely unlike them. It reminds me of how TV shows often have a president that resembles the current president in superficial ways, while being distinct enough that they won't get sued. I'd be interested to know why this happens.
- Kaibeezy 4y agoThere’s a high standard of harm for libel of public figures — knowingly false plus actual malice, iirc. Seems more likely it’s just bad at these details. How would you test it, since it’s not going to have sufficient or reliable source data for ordinary people? ETA: OK, well, based on the following comments, it has a prohibition on living people, but you can’t libel the dead. So it either is bad at faces or it has a prohibition there too. The article would have mentioned if DALL-E said it wouldn’t render Lennon or Harrison. QED, bad at faces?
- CamperBob2 4y agoIt will yell at you and threaten your account with termination if you try to create anything based on a living person's face, from what I can tell.
- j79 4y agoYep. I once tried to create a cartoon dinosaur with hair in the “style” of an ex President (yellow and combed forward), and was warned with a potential ban.
- phonescreen_man 4y agoInterestingly related, I just used AI image generation to create my EP cover.. first I tried running luciddrains dall e 2 PyTorch implementation using the prompt “death by Tetris EP album cover 2022” unfortunately I am using a Mac Pro so the gpu was not able to work. Then I tried imagen PyTorch implementation and used same keyword. This time it was working with the CPU unfortunately 2 days in we had a power outage so I had something but nothing complete. So I fed the generated image into the google dream generator and got my album cover! https://willsimpson.hearnow.com/ https://willsimpson.hearnow.com/
- w0mbat 4y agoHow do you know that the album covers are not part of the corpus of images that DALL-E was trained on in the first place?
- tsimionescu 4y agoIt's interesting that the prompts that would do badly in a Google image search also seem to be the ones that make poor prompts. Basically, it seems that rather than describing a scene, you have to try to give an analogy for some image(s) that it might have in its training set - which is why, I believe, "banana in the style of Andy Warhol" produces a much higher quality result than "Outline of prism on a black background in the middle of scene splits a beam of light coming from the left side into rainbow on the right side".
- alisonkisk 4y ago
- dsign 4y agoIt's going to leave all those artists without a job, you just wait!!
- NonNefarious 4y agoWent to use my invite, and OpenAI demands your PHONE NUMBER. No excuse for it. Screw that.
- pjgalbraith 4y agoI've been recreating the 50 worst heavy metal album art using AI as well, currently at 30. Recently I've found Stable Diffusion plus DALL-E inpainting to be a good combination. https://twitter.com/P_Galbraith/status/1560469019605344256 https://twitter.com/P_Galbraith/status/1560469019605344256
- _HMCB_ 4y agoVery cool. But it just goes to show the impact of human creativity. The conceptual aspect.
- remote_phone 4y agoIf you gave those same instructions to humans I’m sure the output would be just as varied. I’d be interested to see a comparison between dall-e and humans.
- spike021 4y agoI haven't gotten to try it for myself, but I've read a few of these blogs that take you through generating examples or even look-alikes to older art pieces. It surprisingly reminds me a lot of when I traveled to Japan without knowing really any Japanese. I needed to communicate not only with friends who don't know much English either, but also other people (like restaurant wait staff, train station staff, etc.). I used Google Translate often, but many times I or my friend(s) (or the other people) would need to re-write our statements a few times until the translation result clicked well enough in each other's languages to be understandable.
- waveywaves 4y agoDALL E works really well if you are specific enough. When you don't get the intended result, it helps to identify the element which wasn't generated properly and then improve your description of the same. "Two men, one of whom is on fire, shaking hands on an industrial lot." can be rewritten as, "Two men, shaking hands, standing on an industrial lot. Person on the right is on fire. Camera is 30 metres away." You can go into more specifics of the framing and the angle from which you want the picture to be take. By default, DALL E will give you the most realistic generations to your prompts unless you mention "digital art" or a particular art style. I have gotten the best results when generating art instead of photos.
- fimdomeio 4y agoThere are a lot of articles focusing on how close does DALL-E match some prompt, but I wonder if this is a suboptimal way to explore the medium. What if you can get a lot more out of it by embracing the unexpected responses. Can it be a tool for exploring lateral thinking? You provide a prompt computer responds with images that are a prompt to human. A baby swiming next to a dolar bill outputs a distorted person face inside a dolar bill with some baby features, could be the start to a rabbit hole of prompts and images where you'll end up with something completly different than your initial expectations.