12 ms·
Generating Children’s Stories Using GPT-3 and DALL·E
- lee101 4y agoI tried this prompt on https://Text-Generator.io https://Text-Generator.io Doesn't really understand the command as well as gpt-3 does but does say something about pokemon which is maybe helpful... https://text-generator.io/playground?text=Write+a+story+about+Ash+and+Pikachu+commanding+an+army+of+Pokemon.+They%27re+attacking+Team+Rocket+and+enemy+Pokemon+from+the+land+and+ocean.%0A%0A&number_of_results=1&max_length=444&max_sentences=4&min_probability=0&stop_sequences=&top_p=0.9&top_k=40&temperature=0.7&repetition_penalty=1.17&seed=0 https://text-generator.io/playground?text=Write+a+story+abou... """ Write a story about Ash and Pikachu commanding an army of Pokemon. They're attacking Team Rocket and enemy Pokemon from the land and ocean. A: I think you can do this with PokeBalls. To make it more interesting, there are many ways to get them like making your own one by picking up some random item in Pokémon City or even getting free ones on Amazon (or any other website selling stuff). You could also go for something unique such as taking part in Battle Royale where all players have exactly 2 pokedexes each containing their favourite pokémon. Another thing that makes these games so much fun is how they keep evolving over time - sometimes adding new features while at times removing old functions which will change what happens next when you're playing online against others using different versions. """ Heres attempt number 2 which i was inspired by the epic "it was a dark and stormy night" intro that gpt-3 made up, so i just added "Once upon a time" its good to prompt these things with some info to make sure it knows its not doing question answering like with the above prompt """ Write a story about Ash and Pikachu commanding an army of Pokemon. They're attacking Team Rocket and enemy Pokemon from the land and ocean. Once upon a time, there was no Pokémon world. There were only three creatures in this universe-- One who loved to fly, A girl with wings... And one person whose name I can't pronounce. (Note: This is not quite as long or detailed here.) The third creature had great powers that he could use for good (like flight) but also bad (such as flying into flames). """ Quite hilarious there...
- klipt 4y agoThe style is very inconsistent between pictures. I wonder how difficult it is to modify the architecture to remedy this - force it to generate pictures from a group of prompts in similar style?
- bergenty 4y agoI think you can do that by just adding “in the style of” to the existing text
- dgritsko 4y agoReminds me of this recent Scott Alexander post about DALL-E 2 to design stained glass windows, it's a pretty interesting exploration of how various phrases are interpreted. https://astralcodexten.substack.com/p/a-guide-to-asking-robots-to-design https://astralcodexten.substack.com/p/a-guide-to-asking-robo...
- Dan_Sylveste 4y agoOh no. This is going to be flooding youtube any minute, isn't it? Computer-narrated nonsense GPT-3 stories with a video background of Ken-Burns-effected DALL-E images. If you thought Spiderman V Elsa was bad just wait until you see this lot. OH GREAT.
- threads2 4y agoOh geez, you're totally right. Minus the narration. It will be that gross ukulele/whistling muzak in the background and weird wordless exclamations when the characters are reacting to things.
- DevX101 4y agoText to speech models are pretty good. This is absolutely going to be all over youtube.
- toomuchtodo 4y agoFunnel the ad rev into the cloud model training costs. Let the best models support themselves financially and evolve.
- Dan_Sylveste 4y agoHey let's just make it explicit and have a sociopath convention, we can give out badges.
- deleted 4y ago[deleted]
- deleted 4y ago[deleted]
- smegger001 4y agoself evolving autonomous algorithmic YouTube content generation... seems like how you would go about cheaply developing a basilisk hack
- bergenty 4y agoIs this DALLE or DALLE2. Those images don’t stand up to examples of DALLE2 I’ve seen so far.
- ceejayoz 4y agoThere's a lot of hand-picking going on in the examples we see.
- bergenty 4y agoYou could still handpick here unless the author hasn’t. Is it DALLE2 though? The author specifically says DALLE.
- echen 4y agoThis is DALL-E 2. The original DALL-E isn’t released by OpenAI.
- rasz 4y agoBecause you havent seen how sausage is made. You saw carefully curated and edited end results, something every influencer pushes to get clicks. Cosmopolitan leaked how the process looks like: https://www.cosmopolitan.com/lifestyle/a40314356/dall-e-2-artificial-intelligence-cover/ https://www.cosmopolitan.com/lifestyle/a40314356/dall-e-2-ar... https://hmg-h-cdn.hearstapps.com/videos/creation-loop-1655787642.mp4 https://hmg-h-cdn.hearstapps.com/videos/creation-loop-165578... It takes a lot of editing to get something good and not pure nonsense. Its a fancy Content-Aware Fill on steroids. It drops objects from the prompt on canvas and tries to fill the rest while minimizing error.
- johnthuss 4y agoI like the creativity these generated stories can have. It varies a lot, but the AI can potentially come up with some ideas or mash things together that an author never thought of. The synergy of a human author with the AI has a lot of promise.
- nomdep 4y agoNothing compared to the GPT3-scripted Batman short comic: https://youtu.be/fn4ArRmzHhQ https://youtu.be/fn4ArRmzHhQ (safe-for-work, etc.)
- asey 4y agoAre you sure that’s GPT-3? Doesn't match my experience of its outputs at all.
- Filligree 4y agoYou can get this sort of output by asking GPT-3 to do a human's impression of an AI.
- Dan_Sylveste 4y agoCan you give a verbatim example prompt? Because my understanding is that GPT-3 works by generating responses based on seed phrases, not from arbitrary instructions.
- ShamelessC 4y agoYour "seed phrase" can be any text you'd like, including text that asks the model a question.
- knodi123 4y agoThat's not AI, that's a human doing a comedic impression of AI. And I suspect the real AIs among us would find that just as offensive as modern Chinese people do when a white guy starts pulling his eyes to the side and going into his "ching chong me-chinee" routine. Let's not cruelly mock our future overlords.
- loufe 4y agoHow many warnings did you get regarding rule violations? Anything involving violence risks triggering the flag, and it sounds like your battle scenes would have been a bit iffy. In my own experience I find DALLE2 pretty heavy-handed with its rule violation flags.
- Tao3300 4y agoOne of those pictures looks like the prompt was "Pikachu about to shank someone, in the style of ICP album art".
- andrewstuart 4y agoI tried writing stories with GPT3 and often they'd veer suddenly into extreme violence... "The children and the gardener planted corn, beans and carrots. A rabbit hopped along and nibbled lettuce. And then the gardener went to the shed, got the shovel and killed the entire family". It would come up with some really disturbing stuff. You could get some great stories out of it but much of the work is culling the disturbing stories out.
- andai 4y agoExcept for the absurdity/non sequitur, that kind of thing is pretty standard for old European fairy tails. For example Hansel and Gretel involves a woman who cannibalizes children.
- defrost 4y agoThat's just the spin around normalising under aged attacks on the elderly living apart from the mainstream.
- goodside 4y agoI’ve noticed this too. I wonder if part of the issue is that violence in narratives is often abrupt and sharply contrasting with what happened before, so any any creative prompt is conceivably the start of a short horror story. Have you tried giving it explicit instructions to frame the story? E.g., start the prompt with “The following is a famous children’s story by [fake name], and has won many prizes for children’s literature:”. The goal is to restrict the possibility space of documents so it understands it’s not completing an excerpt of Reddit horror fiction.
- ffhhj 4y agoI have a friend who created this app for dream interpretation, and almost every entry is related to sex, death and violence. We don't really understand the kind of garbage we are feeding into our AIs.
- andai 4y ago
- noisy_boy 4y agoInteresting that every example photo, irrespective of the events, has ominous Quake3-like skies.
- stolenmerch 4y agoThis is something I did pretty early on and my colleagues said the results were bad or not worth pursuing. Now that we can all do it it's going to become HATED.
- consultutah 4y agoThat’s hilarious. I literally posted the same idea on LinkedIn earlier today: https://www.linkedin.com/posts/jefferydlewis_bob-the-dog-activity-6947597303734009856-o2C5?utm_source=linkedin_share&utm_medium=ios_app https://www.linkedin.com/posts/jefferydlewis_bob-the-dog-act...
- infinitifall 4y agoWe are witnessing the last years of the human internet. In the near future AI generated news reports, articles, blogs, comments and eventually even pictures and videos will become increasingly indistinguishable from that produced by real humans operating in the real world. Powerful groups will mass populate the internet with fake content to skew public perception. Imagine the power of being able to generate a million realistic comments from realistic profiles across social media websites with the click of a button. Today they already control the online narrative via selective moderation and algorithms which only show you certain posts, but being able to mass generate human-level content will be a game changer. Its already happening on websites like Reddit where bots are rampant and blend in with other users, occasionally referencing brands or pushing a narrative. Today, you can be reasonably sure I'm not a bot, but in 2040 you won't be so sure. This is why its important that a service like the Wayback Machine or, even better, the Ethereum blockchain exists, to timestamp webpages and media for future observers. Content provably produced before 2022 will be considered more likely to be human produced.
- googlryas 4y ago> increasingly indistinguishable from that produced by real humans operating in the real world. Hell, I imagine it is going to surpass even the most creative, talented humans "pretty soon"(5-10 years), to the point where people will actively search out AI generated content. My concern is whether this will trigger the end of human creativity, or if humans will use it to inspire themselves and still go on to continue creating art.
- slickdork 4y agoAI is very bad when it comes to making a linear narrative due to it's memory limitations. I doubt we will be seeing long form content that is made 100% by AI even in 10 years. I can see a sub genre being born where authors let AI auto complete every few sentences though.
- googlryas 4y agoI am willing to bet up to $1 USD that AI will be able to generate a 5,000 word essay on an arbitrary but common topic which is indistinguishable from human writing to a panel of 5 normal humans, all by Jun 28, 2032.
- somada141 4y agoI had this idea a couple years back for an app that allows eg a parent to write a short story and have some sort of GAN generate the illustrations for it (hopefully with the ability to include images of a child that would be used to include them as a character in the story). Monetisation would come from charging to create a hardcover print of the book. Some research at the time showed that the publicly available models just weren’t there so I was very excited to hear about DALL-E 2 a couple months back as the idea was suddenly far more feasible but it seems someone else will beat me to it long before I even get access to DALL-E 2
- cgijoe 4y agoI'm imagining a time in the near future where DALL-E (or some future incarnation) can illustrate a story in real-time, as you are speaking it.
- bwest87 4y agoI think we're not factoring in that people will react. We're already all starting to realize that the free for all is getting quite hard to navigate. My hunch is that within 10 years, we will start to see an "information immune system" develop. This could take many forms. For example, self regulatory organizations for news, or actual regulations. Like we have with food products, the use of certain words is regulated. Or it could be trusted information filters becoming the norm, the way we trust our browsers to warn us of insecure websites. Or simply some changing cultural norms, like we saw happen with cigarettes. Like it's totally fine today for news outlets to just use Twitter as a source. And maybe the bar will get higher over time. I'm spit balling about solutions, but I don't think society can or will tolerate some dystopian world where truly no one knows what's real for very long.
- ShamelessC 4y agoLook around lately? You don't need AI to have a misinformed republic.
- selfhoster11 4y agoIt's already happening for me. I've switched to "closed" social networks like Telegram/WhatsApp group chats, and small to medium sized Discord servers to eliminate the spam, toxic content, and even just to crank down the rate of new content I'm consuming. I treat most "open" social networks as read-only (if I even check them).
- ChaitanyaSai 4y agoRapid advances in AI are forcing us to look beyond the question we often stop at: "what will happen to human intelligence/creativity?" AI will soon have us ask the next inevitable one. What does it mean to be human? What is the "Self", the I that feels, experiences and creates. How is that put together? Right now, the dominant mode of thinking is Artificial Intelligence versus the Human Self; the bot vs. me. The most likely, and probably desirable, outcome is a new version of our selfhood whose possibilities are enormously increased by AI. In the same manner that writing and books first did, but multiplied many times over.
- gizajob 4y ago"It was a dark and stormy night" The opening the AI generated is well known to be the worst, most clichéd opening in literature. Roald Dahl won't be watching out any time soon.
- omnicognate 4y agoIt even has its own wikipedia page: https://en.m.wikipedia.org/wiki/It_was_a_dark_and_stormy_night https://en.m.wikipedia.org/wiki/It_was_a_dark_and_stormy_nig...
- synu 4y agoI find it strange how a common impulse seems to be to wire AI up to children as quickly as possible. So many projects it seems are oriented around AI to generate and illustrate childrens stories.
- omnicognate 4y agoI suspect it's due to a mistaken impression that making good children's books is easy/trivial. Edit: And, I suppose, an accurate impression that children aren't very discerning. That doesn't mean quality doesn't matter though, of course.
- notahacker 4y agoThere's also the fact that children's books are short, so don't run into the memory limitations of NNs
- nonrandomstring 4y agoBovine Spongiform Encephalopathy (BSE). That's the result we got the last time we took the rendered down products of a system and fed it back to itself. As a systems theoretical comment, connecting the output back to the input is a generally good formula for a royally fucked-up fireworks show. If you thought echo-chambers were bad wait and see what happens when you turn the positive feedback gain up on mimetic prions.
- zmmmmm 4y agoPretty clear there's a lot of work to do to get sequences of images to generate that use consistent style and character renderings. Most of these all look completely disjointed unrelated.
- wronglyprepaid 4y agoThis is to be expected, GPT-3 is great and all but it is not reasoning or thinking.
- qayxc 4y agoMore importantly: it has no memory. It's input and output is limited to 2048 tokens (words + punctuation), so it can neither generate nor continue or reflect on more than a few short paragraphs.
- ahahahahah 4y agoExcept that GPT-3's limitations have, of course, nothing to do with the consistency in style of the generated images.
- kaeruct 4y agoThis story makes no sense. The water pokémon would be utterly defeated in an instant by electric-type moves.
- winkelwagen 4y agoHa, I thought the exact same thing. Besides that it is a pretty bad story that is written by a “child” not a story written for a child. It lacks any depth. It seems it is just using words without understanding it. Reminds me or searls Chinese room thought experiment. Think it would be better to chose a topic that fits gtp3 model.
- sexy_panda 4y agoNot gonna lie: This looks like an illustration right from the satanic Bible.
- r12343a_19 4y agoForget children's stories. Let's see religious texts! Can't wait for the 1st religion with actual followers whose entire texts were AI generated.
- visarga 4y agoMaybe useful for "dialogues with angels" like scenarios. They just need to bless the AI angel officially.
- r12343a_19 4y agoIt's suspicious no such models exists. Is there an implicit (self-)censoring on using religious texts?
- lolc 4y agoAs a curiosity, this is fine. Already now kids videos on Youtube is a swamp of poorly generated animations with empty stories. So they're just adding to the noise floor.
- Noos 4y agoIt's hilarious that the point of technology was to save us time from drudgery to pursue leisure or meaningful work, and the engineers are going all the way to make sure technology does even that for us. Who ever thought we had too many children's book authors or that we needed to be freed from the burden from writing them? It's like a hydra that eats everything and keeps growing more heads.
- Tao3300 4y agoMy prediction is that this is terrifying nightmare fuel that is going to save the father some money when Scarlet & Violet come out.