27 ms·
Veo 3 and Imagen 4, and a new tool for filmmaking called Flow
- deleted 1y ago[deleted]
- wingspar 1y agoSo what the copyright situation going to be in an ai generated movie? My last recollection is recent case said AI generated didn’t have copyright?
- BeFlatXIII 1y agoI hope no copyright. Ideas are meant to be freely copied.
- deleted 1y ago[deleted]
- bowsamic 1y agoI'm surprised at how bad these are
- IncreasePosts 1y agoI don't care about AI animals but the old salt offended me.
- Animats 1y agoThe ad for Flow would be much better if they laid off the swirly and wavy effects, and focused on realism. Soon, you should be able to put in a screenplay and a cast, and get a movie out. Then, "Google Sequels" - generates a sequel for any movie.
- colesantiago 1y agoDefinately plausible. All this is in line with my prediction for the first entirely AI generated film (with Sora or other AI video tools) to win an Oscar being less than 5 years away. And we're only 5 months in. https://news.ycombinator.com/item?id=42368951 https://news.ycombinator.com/item?id=42368951
- zanellato19 1y agoI would believe that an AI generated film will never win an Oscar. I bet they will soon add rules that AI movies can't even compete on it.
- zombiwoof 1y agoSeparate categories
- wongarsu 1y agoOscar in what category? We are about six years into transformer models. By now we can get transformers to write coherent short stories, and you can get to novel lengths with very careful iterative prompting (e.g. let the AI generate an outline, then chapter summaries, consistency notes, world building, then generate the chapters). But to get anything approaching a good story you still need a lot of manual intervention at all steps of the process. LLMs go off the rail, get pacing completely wrong and demonstrate gaping holes in their understanding of the real world. Progress on new models is mostly focused in other directions, with better storytelling a byproduct. I doubt we get to "best screenplay" level of writing in five years. Best Actor/Actress/Director/etc are obviously out for an AI production since those roles simply do not exist. Similar with Best Visual Effects, I doubt AI generated films qualify. That leaves us with categories that rate the whole movie (Best Picture, Best International Feature Film etc), sound-related categories (Best Original Score, Original Song, Sound) and maybe Best Cinematography. I doubt the first category is in reach. Video Generation will be good enough in five years. But editing? Screenwriting? Sound Design? My bet would be on the first AI-related Oscar to be for an AI generated original score or original song, and that no other AI wins Oscars within five years. Unless we go by a much wider definition of "entirely AI generated" that would allow significant human intervention and supervision. But the more humans are involved the less it has any claim to being "entirely AI". Most AI-generated trailers or the Balenciaga-Potter-style videos still require a lot of human work
- rxtexit 1y agoI just think the entire framing is wrong. I have done quite a bit with AI generated audio/sound/music. At some point in the process, the end result feels like your own and the models were used to create material for the end work. At some point, using AI in the creative process will be such a given that it is left unsaid. I would assume the screen play next year that wins the Oscar will have been helped with the aid of a language model. I can't imagine a writer not using a language model to riff on ideas. The delusional idea here is the prompt "write an Oscar winning screenplay" and that somehow that is all there is going to the creative process.
- FirmwareBurner 1y ago>Soon, you should be able to put in a screenplay and a cast, and get a movie out. This "fixes" Hollywood's biggest "issues". No more highly paid actors demanding 50 million to appear in your movie, no more pretentious movie stars causing dramas and controversies, no more workers' unions or strikes, but all gains being funneled directly to shareholders. The VFX industry being turned into a gig meatgrinder was already the canary in the coal mine for this shift. Most of the major Hollywood productions from the last 10 years have been nothing but creatively bankrupt sequels, prequels, spinoffs and remakes, all rehashed from previous IP anyway, so how much worse than this can AI do, since it's clear they're not interested in creativity anyway? Hell, it might even be an improvement than what they're making today, and at much lower cost to boot. So why wouldn't they adopt it? From the bean counter MBA perspective it makes perfect sense.
- com2kid 1y ago> Hollywood's wet dream. Except it bankrupts Hollywood, they are no longer needed. Of people can generate full movies at home, there is no more Hollywood. The end game is endless ultra personalized content beamed into people's heads every free waking hour of the day. Hollywood is irrelevant in that future.
- FirmwareBurner 1y agoGood point, this is indeed a threat to them. Like how many young people are watching streamers now instead of worshiping present day's music, TV or movie star like in the 90's. The likes of Youtube and Twitch could be more valuable than Hollywood. That's why I think Hollywood is rushing to adopt gen-AI, so they can churn out personalized content faster and cheaper straight to streaming, at the same rate as indie producers.
- jsheard 1y ago> Of people can generate full movies at home, there is no more Hollywood. LLMs have been in the oven for years longer than this, and I'm not seeing any signs of people generating their own novels at home. Well, besides the get-rich-quick grifters spamming the Kindle store with incoherent slop in the hopes they can trick someone into parting with a dollar before they realize they've been had.
- dimal 1y agoThe swirly effects are probably used to distract from the problems of getting realism right.
- suddenlybananas 1y agoGenerating banal stock footage is wildly different than generating a film.
- bilbo0s 1y agoThis doesn't necessarily preclude the possibility of making a model that can generate a film. It's still something they can work out. In fact, I wouldn't be surprised if models we're seeing these days are not a necessary first step in that process.
- suddenlybananas 1y agoI'm not saying it's in principle impossible, but rather I'm saying this doesn't show that it will happen soon.
- esafak 1y agoAI trailers already exist: https://www.youtube.com/playlist?list=PL_52fVxPZcIiEvGocuVn6fOhbog0Ro8w2 https://www.youtube.com/playlist?list=PL_52fVxPZcIiEvGocuVn6...
- jsheard 1y agoI feel like we should probably draw a distinction between "AI trailers exist as a replacement for traditional trailers" and "AI trailers exist because they're the clickbait du-jour for cynical social media engagement farmers". For now they're 100% the latter.
- elzbardico 1y agoGot a bit of an uncanny valley feeling with the owl and the old man videos. And the origami video give me a sort of sinister feeling, seemed vaguely threatening, agressive.
- vjerancrnjak 1y agoIt's a reflection of yourself. Origami for me was more audio than video. Felt like it's exactly how it would sound.
- jjcm 1y agoLower on the page there's a knitted characters version that feels much better. It seems like for some of these, divorcing yourself from reality a little bit helps avoid the uncanny valley.
- thinkingtoilet 1y agoThe owl one had that glow that so many AI images have for some reason. The man was very impressive to me.
- benlivengood 1y agoWe've made so much progress in the last 20 years; it used to take huge teams of developers and artists and giant compute clusters and rendering time to generate uncanny valley! Now it just takes giant compute clusters and inference time.
- phh 1y agoOf course they had to name a film making proprietary tool with the name of an award winning film made using open-source tools released less than a year ago...
- deleted 1y ago[deleted]
- paxys 1y ago"Flow" is one of the most generic names in tech. I can think of 10+ products called that off the top of my head.
- debugnik 1y agoThere's no way they named their AI filmmaking tool after the last winner of the Academy Award for Best Animated Feature by accident.
- debugnik 1y agoI still remember a style transfer paper which proudly mimicked a popular artist who had passed away barely a few years before (Qinni). Many AI researchers seemingly want to wear the skins of the people they rip off.
- woah 1y agoSeems pretty obvious that they named it after Facebook's JS type checker from 2015
- jonahx 1y ago[flagged]
- deleted 1y ago[deleted]
- quantumHazer 1y agoLike most AI image or video generation tools, they produce results that look good at first glance, but the more you watch, the more flaws and sloppiness you notice, and they really lack storytelling
- superb_dev 1y agoYou don’t even have to look close for some of these. The owl suddenly flipping direction in the first video was jarring
- Workaccount2 1y agoIt doesn't flip, it's an illusion. The owl is always facing the camera.
- quantumHazer 1y agoYeah the owl ""animation"" is terrible, I bet they could have found better examples? If it wasn't the case I don't know what to think
- JamesBarney 1y agoIt looked to me like the owl was turning around to land.
- billyp-rva 1y agoWhen it's in silhouette you don't know what direction it is facing, technically. I think what's happening is when you see a shot of something flying in front of something prominent (in this case, the moon), your brain naturally perceives it is going away from the camera and toward the object.
- llm_nerd 1y agoI think that's just the silhouette illusion[1]. In this case likely abetted by the framing elements moving near the edges. [1] - https://en.wikipedia.org/wiki/Spinning_dancer https://en.wikipedia.org/wiki/Spinning_dancer
- 1y ago
- deleted 1y ago[deleted]
- deleted 1y ago[deleted]
- sergiotapia 1y agoHow do you use Imagen 4 in Gemini? I don't see it in the model picker, I just 2.5 Flash and 2.5 Pro (Upgrade).
- vunderba 1y agoAfter doing some testing, Imagen 4 doesn't score any higher than Imagen 3 on my comparison chart, approximately ~60% prompt adherence accuracy. https://genai-showdown.specr.net https://genai-showdown.specr.net
- Onavo 1y agoHow do companies like https://icon.com https://icon.com do their image Gen if the existing SOTA for prompt adherence is so poor?
- peab 1y agofine tuning and prompt techniques can go a long way. That + cherrypicking results
- yorwba 1y agoPeople who generate images for ads probably don't often need strict prompt adherence, just a random backdrop to slap a picture of their product on top of. The kind of thing they'd have used a stock image library for before. Also "create static + video ads that are 0-99% complete" suggests the performance is hit or miss.
- AsmodiusVI 1y agoExactly this. It just helps the foundation which doesn’t need specific details in most cases.
- htrp 1y agomultishot generation with discriminators
- xixixao 1y agoAwesome showcase! Fun descriptions. Are there similar sites?
- vunderba 1y ago
- carlosdp 1y agoWow, this is incredible work! Blown away at how well the audio/video matches up, and the dialogue is better sounding / on-par with dedicated voice models.
- airstrike 1y agoOn a technical level, this is a great achievement. On a more societal level, I'm not sure continuously diminishing costs for producing AI slop is a net benefit to humanity. I think this whole thing parallels some of the social media pros and cons. We gained the chance to reconnect with long lost friends—from whom we probably drifted apart for real reasons, consciously or not—at the cost of letting the general level of discourse to tank to its current state thanks to engagement-maximizing algorithms.
- pelagicAustral 1y agoHave they reveled anything similar to Claude Code yet? I sure hope they are saving that for I/O next month... this video/photo reveals are too gimmicky for my liking, alas I'm probably biased because I don't really have a use for them.
- dmd 1y agohttps://jules.google/ https://jules.google/ posted here today https://news.ycombinator.com/item?id=44034918 https://news.ycombinator.com/item?id=44034918
- pelagicAustral 1y agoYeah, I saw that... not quite the same... I used it for a bit but it's more like an agent that clings to a Github repo and deals with tickets up there, can't really test live on local, it just serves a different purpose.
- lxgr 1y agoGoogle I/O is happening right now. This is one of the announcements, I believe.
- jader201 1y agoI'm surprised no one has yet to mention the use of the name "Flow", which is also the title of the 2025 Oscar winning animated movie, built using Blender. [1] This naming seems very confusing, as I originally thought there must be some connection. But I don't think there is. [1] https://news.ycombinator.com/item?id=43237273 https://news.ycombinator.com/item?id=43237273
- imp0cat 1y agoWould Google really stoop so low and try to use the success of the movie to prop their AI video generator tool? But then again, the do no evil motto is long gone, so I guess anything goes now?
- Legend2440 1y agoIt's a common word. There are like 50 things named Flow. It's unrelated.
- a2128 1y agoBefore picking a name it's necessary to Google it and make sure you're not squatting on anything important. It's hard to believe that they didn't find they're about to squat on an Oscar winning animated film less than a year after its release. They decided to roll with it anyway, for a tool that basically aims to eliminate animators and filmmakers
- lnyan 1y agoNote that it's very likely that Veo models are based on "Flow Matching" [1] [1] https://arxiv.org/abs/2210.02747 https://arxiv.org/abs/2210.02747
- maldie 1y agoFor sure seems they are likely deliberately riding on the fame of the movie. I too instantly thought it is some kind of Flow movie animation collaboration similarily like Flow is represented in Blender 4.4 splash screen or is even their mascot.
- Workaccount2 1y agoI'm sure by this point, and if not, pretty soon, everyone will have seen a clip of AI generated video and not thought twice about it. Its something that is only obvious when it is obvious. And the more obvious examples you see, the more non-obvious examples slip by.
- gpt5 1y agoI saw a video today [1]. Millions of views, ten thousand comments, not a single commenter mentioned that it's AI generated. If you look at the shadows in the background, you can see how they appear and disappear, how things float in the air, and have all the AI artifacts. The video is also slowed down (lower FPS) to overcome the length limit of AI video generator. But the point is not how we can spot these, because it's going to be impossible, but how the future of news consumption is going to look like. [1] https://www.tiktok.com/@calm.with.word/video/7505837083274120478 https://www.tiktok.com/@calm.with.word/video/750583708327412...
- xarope 1y agothe detail around the eyes is a dead-giveaway for AI generated video
- jadamson 1y agoInteresting. Here's the original channel - their videos all have a Picsart watermark in the bottom right. I don't believe it's entirely fake, just enhanced. https://www.youtube.com/@jrcollection5246/shorts https://www.youtube.com/@jrcollection5246/shorts
- kilpikaarna 1y agoWell, what does news (or any media) consumption look like now? It's been trending towards pure noise for a good while, and this is a way to further automate the generation of yet more noise.
- RobertBobert 1y ago[dead]
- pier25 1y agowhat do they use to train these models? youtube videos?
- jonplackett 1y agoHas anyone actually tried Veo3 and know if it’s as good as this looks? The demo videos for Sora look amazing but using it is substantially more frustrating and hit and miss.
- gpt5 1y agoHere is a twitter user that is posting videos generated with Veo3 (watch unmuted): https://x.com/fofrAI https://x.com/fofrAI
- sebau 1y agoFuture is not bright. While we are endlessly talking about details reality is that AI is taken over so many jobs. Not in 10 years but now. People who just see this as terrible are wrong. AI improving curves is exponential. People adaptability is at best linear. This makes me really sad. For creativity. For people.
- mindvirus 1y agoMaybe. The internet was also exponential, and while it has its drawbacks, I think it's resulted in a huge increase in creativity. The world looks very different than it did 30 years ago, and I think mostly for the better.
- jampekka 1y ago> Future is not bright. While we are endlessly talking about details reality is that AI is taken over so many jobs. Of course this is not because of AI. It's because of the ridiculous system of social organization where increased automation and efficiency makes people worse off.
- elzbardico 1y agoTime for the Butlerian Jihad
- jjcm 1y agoIt finally feels like the professional tools have greatly outpaced the open source versions. While wan and hunyuan are solid free options, the latest from Google and Runway have started to feel like a league above. Interestingly it feels like the biggest differentiator is editing tools - ability to prompt motion, direction, cuts, or weaving in audio, rather than just pure ability to one shot. These larger companies are clearly going after the agency/hollywood use cases. It'll be fascinating to see when they become the default rather than a niche option - that time seems to be drawing closer faster than anticipated. The results here are great, but they're still one or two generations off.
- javchz 1y agoI think open source still has an important advantage in the pro environment despite being less convenient, and it's the possibility of adding things in between the generation process like control net, and custom loras with new concepts or characters. Plus in local generation you're not limited by the platform moderation that can be too strict and arbitrary and fail with the false positives. Yes comfy UI can be intimidating at first vs an easy to use chatgpt-like ui, but the lack of control make me feel these tools will still not being used in professional productions in the short term, but more in small YouTube channels and smaller productions.
- popalchemist 1y agoControl net etc can be served via API; the intrinsic advantage of open-source is the ability to train and run inference privately.
- deleted 1y ago[deleted]
- doctorpangloss 1y agoSomeone out there might care about nudity, but unfortunately, nobody that matters.
- MrScruff 1y ago
- nrjames 1y agoThis is technically impressive and I commend the team that brought it to life. It makes me sad, though. I wish we were pushing AI more to automate non-creative work and not burying the creatives among us in a pile of AI generated content.
- yieldcrv 1y agoI’m a creative and I’m really glad that more people can express themselves Just wanted to add representation to that feeling
- lilwobbles 1y agoExpressing themselves by generating boilerplate content? Creativity is a conversation with yourself and God. Stripping away the struggle that comes with creativity defeats the entire purpose. Making it easier to make content is good for capital, but no one will ever get fulfillment out of prompting an AI and settling with the result.
- ivape 1y agoCheck out all the creatives on /r/screenwriting, half the time they are trying to figure out how to "make connections" just to get a story considered. It's a fucking nightmare out there. Whatever god is providing us with AI is the greatest gift I could imagine to a creative.
- onemoresoop 1y agoAI could be useful if used like any other tool, but not as an all in box where everything is done for you minus the prompt. Im actually worried people will become lazy
- deleted 1y ago[deleted]
- drusepth 1y ago
- rvz 1y agoWell, all the AI labs wanted to "Feel the AGI" and the smoke from Google... They all got smoked by Google with what they just announced.
- seydor 1y ago[flagged]
- deleted 1y ago[deleted]
- ugh123 1y agoWhen can I change the camera view and have everything stay consistent?
- skybrian 1y agoWhat’s the easiest way to try out Imagen 4? Edit: https://labs.google/fx/tools/whisk https://labs.google/fx/tools/whisk
- flakiness 1y agoHow does this compare with sora (pro)?
- echelon 1y agoSora, the video model, is shit. Kling, Runway, and a whole host of other models are better. You don't have to do much to be better than Sora. Sora, the image model (gpt-image-1), is phenomenal and is the best-in-class. I can't wait to see where the new Imagen and Veo stack up.
- 999900000999 1y agoEhh, really for 20$. Break dancers with no music, people just pop in and out ? Google what is this? How would anyone use this for a commercial application.
- lenerdenator 1y agoI do find myself wondering if the people working on this stuff ever give any real thought to the impact on society that this is going to have. I mean obviously the answer is "no" and this is going to get a bunch of replies saying that inventors are not to blame but the negative results of a technology like this are fairly obvious. We had a movie two years ago about a blubbering scientist who blatantly ignored that to the detriment of his own mental health.
- tmpz22 1y agoHow could you possibly push back on the societal benefit of a director being able buy a vacation home in Lake Tahoe?
- lamp_book 1y agoAnd what about the rest of human pyramid working under the director employed in these productions?
- bowsamic 1y agoIt's really being forced on us too. Jira, Confluence, and Notion are three products I've used where they've purposefully ignored requests to allow us to disable or hide the bundled generative AI. It's really intrusive. I also switched to Duck Duck Go because of the new AI on Google
- tootie 1y agoRemember when they fired Timnit Gebru for publishing on AI safety?
- themacguffinman 1y agoQuite a narrow view to interpret what happened there as firing Gebru for publishing on AI safety. Google still conducts and publishes research on AI safety, just without Gebru who helpfully offered to resign if Google didn't name her critics.
- dragonwriter 1y ago
- StefanBatory 1y agoThanks to them, we will be able to enter new era of politics. Where nothing is true, and everything is vibe based. Thank you, researchers, for making our world worse. Thank you for helping to kill democracy.
- matthewaveryusa 1y ago"The Bloomberg terminal for creatives"
- crat3r 1y agoThis doesn't look (any?) better than what was shown a year or two ago for the initial Sora release. I imagine video is a far tougher thing to model, but it's kind of weird how all these models are incapable of not looking like AI generated content. They all are smooth and shiny and robotic, year after year its the same. If anything, the earlier generators like that horrifying "Will Smith eating spaghetti" generation from back like three years ago looks LESS robotic than any of the recent floaty clips that are generated now. I'm sure it will get better, whatever, but unlike the goal of LLMs for code/writing where the primary concern is how correct the output is, video won't be accepted as easily without it NOT looking like AI. I am starting to wonder if thats even possible since these are effectively making composite guesses based on training data and the outputs do ultimately look similar to those "Here is what the average American's face looks like, based on 1000 people's faces super-imposed onto each other" that used to show up on Reddit all the time. Uncanny, soft, and not particularly interesting.
- ahmedfromtunis 1y agoIt has long been established that Veo has a waaay better understanding of physics, and consistency over multiple frames, than Sora. Not even close.
- crat3r 1y agoI want to be clear, I don't think Sora looks better. What I am saying is they both look AI generated to a fault, something I would have thought would be not as prominent at this point. I don't follow the video generation stuff, so the last time I saw AI video it was the initial Sora release, and I just went back to that press release and I still maintain that this does not seem like the type of leap I would have expected. We see pretty massive upgrades every release between all the major LLM models for code/reasoning, but I was kind of shocked to see that the video output seems stuck in late 2023/early 2024 which was impressive then but a lot less impressive a year out I guess.
- htrp 1y agois it still a waitlist?
- Imnimo 1y ago>Imagen 4 is available today in the Gemini app, Whisk, Vertex AI and across Slides, Vids, Docs and more in Workspace. I'm always hesitant with rollouts like this. If I go to one of these, there's no indication which Imagen version I'm getting results from. If I get an output that's underwhelming, how do I know whether it's the new model or if the rollout hasn't reached me yet?
- minimaxir 1y agoGoogle is typically upfront about which model versions you're using in those tools. Not as behind-the-scenes as ChatGPT. However, looking at the UI/UX in Google Docs, it's less transparent.
- cubefox 1y agoIndeed. At the bottom of their Imagen page, they link to Google AI Studio: https://aistudio.google.com/generate-image https://aistudio.google.com/generate-image But this still says it's Imagen 3.0-002, not Imagen 4.
- matsemann 1y agoYes, Google is so, so, so bad at this. I even struggle with gemini often telling me it can't make images, until I tell it that it can, and then it does. I have no idea what's really supposed to be supported or not in gemini. It is so confusing. Ok, I got gemini pro through workspace or something, but not everything is there? Sure, I can try aistudio, flow, veo, gemini etc to figure out what I can do where, but so bad UX. Just tried using gemini to create an image, definitely not the newest imagegen as the text was just marbled up. But I can't see which version I'm on, genious. Edit: After clicking through lots of google products I'm still not able to find a single place I can actually try the new imagegen, despite the article claiming it's available today in X,Y,Z
- nico 1y agoWow, the audio integrations really makes a huge difference, especially given it does both sounds and voices Can’t wait to see what people start making with these
- gloosx 1y ago>>models create, empowering artists to bring their creative vision Interesting logic the new era brings: something else creates, and you only "bring your vision to life", but what it means is left for readers questioning, your "vision" here is your text prompt? Were at a crossroads where the tools are powerful enough to make the process optional. That raises uncomfortable questions: if you don’t have to create anymore, will people still value the journey? Will vision alone be enough? What's the creative purpose in life? To create, or to to bring creative vision to life? Isn't the act of creation is being subtly redefined?
- dmonitor 1y agoIt's being redefined in such a way that 2-3 very large entities get to hold the means of production. It's a very convenient redefinition for them.
- teitoklien 1y ago[flagged]
- gloosx 1y agoIf your focus is to solve the problem, then it makes sense to treat the process as secondary. The tools are just means to an end. This view also aligns with how generative AI is marketed – it's a way to accelerate realization, not a way to focus on the act of crafting. That said, outcome-first thinking does run the risk of disconnection, and our current culture is all about disconnection.
- teitoklien 1y agoEven the process is more democratized now, Want to learn coding ? Build out your app idea first with replit -> Then export the codebase into your computer -> Run claude code on it and ask it to scan all the files and describe the tech stack to you and how it operates while giving you all the major components you need to learn to understand it with youtube channel and book recommendations for each topic + work exercises -> Use perplexity deep research once a week to further research every topic as you start to learn them If you’re a busy man/woman make gumloop or lindyai workflow to check your calendar and pack in timeslots to do all of this learning, and then auto send you worksheets via email as homework to test you skills All of this for a price of 1/15th of a college degree (not even an expensive college) This is not hypothetical conjecture I do this daily. So everyone has now 1) Low cost access to build stuff with one prompt to realise the value of tools 2) A personal tutor that can then help you scour the depths of the craft and force you to practice and learn deeply now with your added motivation of knowing what’s possible with building stuff So it has the potential to connect us more too, it’s upto humans to choose whether they do at the end tho. That is their liberty.
- julianpye 1y agoAn indie film with poor production values, even bad acting can grip you, make you laugh and make you cry. The consistency of quality is key - even if it is poor. The directing is the red thread throughout the scenes. Anything with different quality levels interrupts your flow and breaks your experience. The problem with AI video content at this stage is that the clips are very good 'in themselves', just as LLM results are, but putting them together to let you engage beyond an individual clip will not be possible for a long time. It will work where the red thread is in the audio (e.g. a title sequence) and you put some clips together to support the thread. But Hollywood has nothing to fear at this stage. In addition, remember that visual artists are control freaks of the purest kind. Film is still used because of the grain, not despite it. 24p prevails.
- doctorpangloss 1y agoThere’s already more good content than anyone can watch. It’s impossible to disentangle strength of the art from strength of distribution. Google, the world’s biggest distributor of culture, is focusing on this problem they do not need to solve, instead of the one everyone in art actually suffers from, because: they’re bad at this. It’s that simple.
- sandspar 1y agoAI video may be to Hollywood as photography was to painting. Photography wasn't "painting, but better" - it was a different thing. AI-native video may not resemble typical Hollywood 3-act structure. But if it takes enough eyeballs away from Hollywood then Hollywood will die all the same.
- pedalpete 1y agoI think you're contradicting your own argument. Painting didn't die from photography. Photography increased the abstract and more creative aspects of painting and created a new style because photography removed much of the need to capture realism. Though, I am still entranced by realist painting style myself, it is serving different purpose than capturing a moment.
- brm 1y agoI think it's a good thing to have more people creating things. I also think it's a good thing to have to do some work and some thinking and planning to produce a work.
- curvaturearth 1y agoThe first video is problematic? the owl faces forwards then seamlessly turns around - something is very off there. The guy in the third video looks like a dressed up Ewan McGregor, anyone else see that? I guess we can welcome even more quality 5 second clips for Shorts and Instagram
- itissid 1y agoWho is doing all the work of making physical agents that can behave as good as a UBI generator? Something that can not just create videos, but go get groceries(hell grow my food), help a construction worker lay down tiling, help a nurse fetch supplies. https://www.figure.ai/ https://www.figure.ai/ does not exist yet, at least not for the masses. Why are Meta and Google just building the next coder and not the next robot? Its because those problem are at the bottom of the economic ladder. But they have the money for it and it would create so much abundance, it would crash the cost of living and free up human labor to imagine and do things more creatively than whatever Veo 4 can ever do.
- pj_mukh 1y agoWelcome to the defining paradox of the 21st century: https://en.wikipedia.org/wiki/Moravec%27s_paradox https://en.wikipedia.org/wiki/Moravec%27s_paradox
- BosunoB 1y agoThere are companies working on this, but my understanding is that the training data is more challenging to get because it involves reinforcement learning in physical space. In the forecast of the AI-2027 guys, robotics come after they've already created superintelligent AI, largely just because it's easier to create the relevant data for thinking than for moving in physical space.
- throwaway314155 1y agoI think I have a similar distaste for Google as you, but it's just due to limitations in the (bleeding edge...) technology. There's not like a conspiracy to _not_ make a "UBI generator" - which is surely not possible with current technology and won't be for awhile however hard Google might try.
- deleted 1y ago[deleted]
- ericskiff 1y agoHas anyone gotten access to Imagen 4 for image editing, inpaint/outpaint or using reference images yet? That's core to my workflow and their docs just lead to a google form. I've submitted but it feels like it's a bit of a black hole.
- cryptoegorophy 1y agoFor anyone with an access, can you ask it to make a pickup truck drive through mud? I’ve tested various different AIs and they all suck with physics and tires spinning wrong way, it is just embarrassing. Demos look amazing, but when it comes to actual use - there is none that worked for me. I guess it is all to increase “investor value”
- roskelld 1y agoGoogle posted a video of their own of an off-roader going through mud. https://www.youtube.com/watch?v=SPF4MGL7K5I https://www.youtube.com/watch?v=SPF4MGL7K5I Obviously we don't know how hand picked that is so it would be interesting to see a comparison from someone with access.
- lelandbatey 1y agoI think Google's got something going wrong with their usage limits, they're warning I'm about to hit my video limit after I gave two prompts. I have a Google AI Pro subscription (came free for 1 year with a phone) and I logged into Flow and provided exactly 2 prompts. Flow generated 2 videos per prompt, for a total of 4 videos, each ~8 seconds long. I then went to the gemini.google.com interface, selected the "Veo 2" model, and am now being told "You can generate 2 more videos today". Since Google seems super cagey about what their exact limits actually are, even for paying customers, it's hard to know if that's an error or not. If it's not an error, if it's intentional, I don't understand how that's at all worth $20 a month. I'm literally trying to use your product Google, why won't you let me?
- kapildev 1y agoGoogle has partnered with Darren Aronofsky’s AI-Driven Studio Primordial Soup. I still don't understand why SAG-AFTRA's strike to ban AI from Hollywood studios didn't affect this new studio. Does anyone know?
- cjkaminski 1y agoPrimordial Soup isn't a guild signatory, which means they aren't bound by the agreement negotiated during the strike. It also means they cannot hire guild actors for their projects, but that isn't a likely concern given the nature of the company.
- ionwake 1y agoLove flow tv ! Absolutely blown away by the improvements on these models, and also the channel interface was not bad and quite smooth. I cant be the only one wondering where the swedish beach volleyball channel is though.
- Lucasoato 1y ago> Flow is not available in your country yet. A bit depressing.
- fasdfdsa 1y ago[flagged]
- fefawfefafds 1y ago[dead]
- fdaffeafe 1y ago[dead]
- kumarm 1y agoAll my Veo 3 videos has sound missing. No idea why. Seems like a common problem.
- merillecuz56 1y ago[dead]
- cynicalpeace 1y agoBasic principles: 1. People like to be entertained. 2. NeuralViz demonstrates AI videos (with a lot of human massaging) can be entertaining To me the fundamental question is- "will AI make videos that are entertaining without human massaging?" This is similar to the idea of "will AI make apps that are useful without human massaging" Or "will AI create ideas that are influential without human massaging" By "no human massaging", I mean completely autonomous. The only prompt being "Create". I am unaware of any idea, app or video to date that has been influential, useful or entertaining without human massaging. That doesn't mean it can't happen. It's fundamentally a technical question. Right now AI is trained on human collected data. So, technically, It's hard for me to imagine it can diverge significantly from what's already been done. I'm willing to be proven wrong. The Christian in me tells me that Humans are able to diverge significantly from what's already been done because each of us are imbibed with a divine spirit that AI does not have. But maybe AI could have some other property that allows it to diverge from its training data.
- methuselah_in 1y agoWell all this is great from a technology point of view. But what about millions of jobs in the film industry in animation, motion artists etc? Why is it feeling like few humans are making sure others stop eating and living a good life?
- codezero 1y agowhat about all the cartographers, the printing press workers, the stables that tend to the working horses? Technology is inevitable and it's a tool, advancing technology will always leave people who specialize and are unable to adapt in a bad position, but this won't stop technology from advancing. I think one could argue this is one of the reasons many people would like their community/government to provide social safety nets for them. It would make specializing less risky in a time when technology advances at a fast pace.
- danabenson 1y agohttps://en.wikipedia.org/wiki/Luddite https://en.wikipedia.org/wiki/Luddite
- sidibe 1y agoThis is coming for everyone's jobs. It'd be possibly an OK or good thing after some adaptation if I didn't suspect that the people with power during this transition were nihilists or people who's mission in life is to be relatively rather than absolutely well off. If everyone can have what they need they will not feel important enough
- jampekka 1y agoYou are probably meaning that as a slur. In reality Luddites did not oppose technology per-se, but the dramatic worsening of the working conditions in the factories, reduced wages and concentration of the income to the capital holders. These are the same problems that should be addressed contemporarily. They initially tried to address these by political means. But with that failing they moved to sabotage and violence. https://www.smithsonianmag.com/innovation/when-robots-take-jobs-remember-luddites-180961423/ https://www.smithsonianmag.com/innovation/when-robots-take-j...
- numpad0 1y agoI came across some online threads sharing LoRA models the other day - and it seemed that a lot of generative AI users seem to share models that are effectively just highly specialized fixed function filters for existing (generated)images? The obvious aim of these foundational image/movie generation AI developments is for these to become the primary source of values at cost and quality unparalleled by preexisting human experts, while allowing but not necessitating further modifications by now heavily commoditized and devalued ex-professional editors at downstream to allow for their slow deprecation. But the opposite seem to be happening: better data are still human generated, generators are increasingly human curated, and are used increasingly closer to the tail end of the pipeline instead of head. Which isn't so threatening nor interesting to me, but I do wonder if that's a safe, let alone expected, outcome for those pushing these developments. Aren't you welding a nozzle onto open can of worms?
- _ncuy 1y agoGoogle hit the jackpot with their acquisition of YouTube and it's now paying dividend. YouTube is the largest single source of data and traffic on the Internet, and it's still growing fast. I think this data will prove incredibly important to robotics as well. It's a shame they sold Boston Dynamics in one of their dumbest ever moves because of bad PR.
- brunoborges 1y ago"Growing fast" is questionable these days. There is an ever growing percentage of new AI-generated videos among every set of daily uploads. How long until more than half of uploads in a day are AI-generated?
- franga2000 1y agoAnd google is in the best possible position to detect it if they want to exclude it from their datasets.
- sebstefan 1y agoThey're never going to manage to do that, just on a technical level Plus some users might want to legitimately upload things with AI-generated content in it
- Timon3 1y ago> They're never going to manage to do that, just on a technical level Why not? Given enough data, it's possible to train models to differentiate - especially since humans can pick up on the difference pretty well. > Plus some users might want to legitimately upload things with AI-generated content in it Excluding videos from training datasets doesn't mean excluding them from Youtube.
- _delirium 1y agoI agree, especially because in practice the vast majority of AI-generated videos uploaded to YouTube are going to be from one of about 3 or 4 generators (Sora, Veo, etc.). May change in the future, but at the moment the detection problem is pretty well constrained.
- celespider 1y agoI have some base knowledge about diffusion/dit, I am so curious about how this can be done. Do you know some resources in this field? THANKS!
- tianshuo 1y agoFeel free to test imagen 4 on this benchmark: https://github.com/tianshuo/Impossible-AIGC-Benchmark https://github.com/tianshuo/Impossible-AIGC-Benchmark Ideogram and gpt4o passes only a few, but not all of them.
- sech8420 1y agoFirst test is... very confusing - https://x.com/Seancheno/status/1925049073230372980 https://x.com/Seancheno/status/1925049073230372980
- baxtr 1y agoA whale coming out of the street in Manhattan, a women with a Jellyfish belly walking in the woods. Why is it that all these AI concept videos are completely crazy?
- rafaelmn 1y agoI'm going to go out on a limb and say because it's easiest to take whatever comes out looking interesting and sell it as a vibe ? Like if you asked a model to help you create a coffeeshop website for a demo, it started looking more like sex shop, you just vibe with it and say that's what you wanted in the first place. I've noticed that the success rate of using AI is proportional to much you can gaslight yourself.
- kypro 1y agoIf the concept is unrealistic your mind will be more forgiving to unrealisms. But if it's suppose to be photo-realistic, you'll be hyper-critical.
- matsemann 1y agoThere is a point in that since you don't know how these really should look you can't really judge them on small idiosyncrasies, and hence you get a better impression compared to uncanny valley if it's something common. However, I also think this is to show that it can create anything, not just copies of stuff it has seen. If you ask for a painting of a woman and it shows you mona lisa, that's not very impressive.
- arduinomancer 1y agoI can definitely see this being used for lower end advertising I’ve noticed ads with AI voices already, but having it lip synced with someone talking in a video really sells it more
- Daub 1y agoAs an artist and designer (with admittedly limited AI experience), where I feel AI to be lacking is in its poverty of support for formal descriptors. Content descriptors such as 'dog wearing a hat' are a mostly solved problem. Support for simple formal descriptors such as basic color terms and background/foreground are ok, but things like 'global contrast' (as opposed to foreground background contrast), 'negative shape', 'overlap', 'saturation contrast' etc etc... all these leave the AI models I have played with scratching their heads. I like how Veo supports camera moves, though I wonder if it clearly recognizes the difference between 'in-camera motion' and 'camera motion' and also things like 'global motion' (e.g. the motion of rain, snow etc). Obligatory link to Every Frame a Painting, where he talks about motion in Kurosawa: https://www.youtube.com/watch?v=doaQC-S8de8 https://www.youtube.com/watch?v=doaQC-S8de8 The abiding issue is that artists (animators, filmmakers etc) have not done an effective job at formalising these attributes or even naming them consistently. Every Frame a Painting does a good job but even he has a tendency to hand wave these attributes.
- dsadfjasdf 1y ago[flagged]
- ssijak 1y agoOlder people on social networks are cooked. I mean in general, we are entering an age where making scams and spreading false news will be easily done with 10$ of credits.
- asl2D 1y agoYeah i fear that too, my grandma is already sending me links of AI animals that she thinks is real, and the horrible/beautiful art of facebook memes/holiday cards, seems to be completely overtaken by AI. We know that full fake video of you with your own voice asking for something or even interacting on a video call is basically solved problem. Prime time to reestablish and confirm trusted channels with the people you care about.
- Workaccount2 1y agoRecently at a family dinner we established that any kind of unsolicited contact that falls outside typical conversation - Asking for money, sending money, pretty much anything with money - you must say the word that only people in our family would know. Ironically, this would be a good application of AI, where the AI listens in on their calls, and will flag conversation that warrants the keyword being said.
- anilgulecha 1y agoI'd made a prediction/bet a month ago, predicting 6 months to a full 90 minute movie by someone sitting on their computer. [0] The pace is so crazy that was an over estimation! I'll probably get done in 2. Wild times. 0: https://www.linkedin.com/feed/update/urn:li:activity:7317975013779808256/ https://www.linkedin.com/feed/update/urn:li:activity:7317975...
- DoesntMatter22 1y agoIt's doable now. Someone just needs to do it. With voice now it's completely doable. Just throw it all together add some effects and you've got a great movie... In theory
- _xhhf 1y agoIt's not a theory, at Cannes a feature movie has premiered that is generated entirely by AI. Made in Spain.
- anilgulecha 1y agocan you share a link/details, please?
- jbkkd 1y agohttps://ground.news/article/the-great-reset-the-first-photorealistic-ai-film-makes-history-at-the-cannes-film-festival_17f758 https://ground.news/article/the-great-reset-the-first-photor...
- DoesntMatter22 1y agoIt's a great movie in theory. Idk how good the movie you mentioned is
- rtkwe 1y agoThere's still a lot of work to be done. It's good at making short individual scenes but when you start trying to string them together the wheels start to come off a lot. This [0] pretty basic police raid leads to shootout video for example turns to mush pretty quick because even in the initial car ride the interior of the car's size and shape warps pretty drastically. Feels like there's going to be a dichotomy where the individual visuals look pretty good taken by themselves but the story told by those shots will still be mushy AI slop for a while. I've seen this kind of mushy consistency hold up over the generations so far, it seems very difficult to remove becasue it relies on more context than just previous images and text descriptions to manage. [0] https://www.reddit.com/r/ChatGPT/comments/1kru6jb/this_video_is_completely_aigenerated_from_video/ https://www.reddit.com/r/ChatGPT/comments/1kru6jb/this_video...
- oliwary 1y agoThis demo video of Veo 3 on reddit, featuring a variety of characters talking in different scenarios and accents, is one of the most incredible AI demos I have ever seen: https://www.reddit.com/r/ChatGPT/comments/1krmsns/wtf_ai_videos_can_have_sound_now_all_from_one/ https://www.reddit.com/r/ChatGPT/comments/1krmsns/wtf_ai_vid... Created by Ari Kuschnir
- marcyb5st 1y agoCalling it now. Someone will use AI to make the "AI Killed the Video Star" video. Probably the same guy that made this[1] and other masterpieces. [1]https://www.youtube.com/watch?v=EICWYazyqu4 https://www.youtube.com/watch?v=EICWYazyqu4
- epiccoleman 1y agoI thought you were going to link "Video Killed The YTMND Star" - which gives me quite the dose of nostalgia: https://www.youtube.com/watch?v=D6D9arrHiLE https://www.youtube.com/watch?v=D6D9arrHiLE
- kridsdale1 1y agoIt’s true. YTMND nailed the TikTok / Vine format like 12 years ahead of its time. If only they’d “pivoted to mobile” and added more ease of use creation tools they may have stayed relevant.
- IanCal 1y agoGood lord. I think the change here will be something we've seen with the other modalities. Text was interestingly syntactically correct but nonsense sentences. Then paragraphs but the end of the article would go off the rails. Then the article. Now it's that the creativity of the children's story in question. Pictures were awful fever dreams filled with eyes but you could kind of see a dog. Then you could see what it was, then decent Videos were fun that they kind of worked, then surprising it took a few seconds for the panda to turn into spaghetti, then it kept the general style for a decent time. I see this moving towards the creativity being the major thing, or it having a few general styles (softly lit background for example). This has mostly all shifted in a very short space of time and as someone who put RBMs on GPUs possibly for the first time (I'm gonna claim it) this is absolutely wild. Had I seen some of this, say, 6 months ago I'd not have guessed at all bits weren't real.
- onlyreal_1 1y agotbh, wasnt that impressed maybe its cause social media has been heavily marketing out all these things in bulkkk and moreover, at this point, it just feels one company copying what the other released, even the names feel not original?
- horhay 1y agoI generally think that Kling or even Runway has achieved the visual fidelity of Veo (flaws and all, physics problems and direction of action and such), but now people are basically experiencing sensory bias where they think that some things about the visuals make better sense because nw it has sound as an added context. Visually, yeah. Probably on par with Kling, possibly worse on the depction of dynamic action
- nprateem 1y agoStability is conspicuously absent from the imagen benchmarks. I assume that means it's significantly better
- TheAceOfHearts 1y agoI tried Whisk to generate images which I then animated, thinking it would be using the newest model. But then I noticed that Veo 3 and Imagegen 4 are only usable through Flow, and only if you're on the most expensive plan. AI Studio also only shows Imagegen3 and Veo2 as media generating options. My main issue when trying out Veo 2 was that it felt very static. A couple elements or details were animated, but it felt unnatural that most elements remained static. The Veo 3 demos lack any examples where various elements are animated into doing different things in the same shot, which suggests that it's not possible. Some of the example videos that I've seen are neat, but a tech demo isn't a product. It would be really cool if Google contracted a bunch of artists / directors to spend like a week trying to make a couple videos or short movies to really showcase the product's functionality. I imagine that they don't do that because it would make the seams and limitations of their models a bit too apparent. Finally, I have to complaint that Flow claims to not be available in Puerto Rico: "Flow is not available in your country yet." Despite being a US territory and being US citizens.
- Workaccount2 1y agoYou can use imagen 4 in vertex ai. But no Veo 3. Also Google is going to have to tread carefully, people in the entertainment industry are already AI hostile, and they dictate a surprising amount of public opinion.
- dsadfjasdf 1y agoYou can use veo 2 for free in the google ai dashboard. like 5 a day
- afroboy 1y agoCan we talk about the elephant in the room, porn and i mean the weird and dangerous one? that moment in history of AI is going to happen and when it did shit will hit the fan.
- asl2D 1y agoDo you think anybody will really care? People were generating CSAM basically as soon as image generation become accessible. And for the less dangerous stuff situation is way more rampant already, both in free and commercial way.
- UncleMeat 1y agoYes. Deepfaked porn is already a widespread mechanism for harassment, both among children and adults. As it gets even easier to create and more and more convincing it will just get worse.
- HamsterDan 1y agoYou should have AI start writing your comments for you so at least then they'll make sense.
- Flamentono2 1y agoAI porn already exist. Im pretty sure kid/child ai porn already exist somewhere. But i'm quite lucky despite knowing rotten.com and plenty of other sides, never having seen real so i doubt i will see fake child porn. Whats the elephant in the room now? Nothing changed. Whoever consumes real will consume fake too. FBI/CIA will still try to destroy cp rings. We could even think it might make this situation somehow better because they might consume purely virtual cp?
- nielsbot 1y agoI've already seen headlines about this: https://arstechnica.com/tech-policy/2025/02/25-arrested-so-far-for-sharing-ai-sex-images-of-minors-in-largest-eu-crackdown/ https://arstechnica.com/tech-policy/2025/02/25-arrested-so-f...
- ravenical 1y agoWhy do people making AI image tools keep showing "pixel art" made with it when the tools are so obviously bad at making it? it's such a basic unforced error
- skc 1y agoI'm excited about this. Think of all of your favorite novels that are deemed "impossible" to adapt to the screen. Or think of all the brilliant ideas for films that are destined to die in the minds of people who will never, ever have the luck or connections required to make it to Hollywood. When this stuff truly matures and gets commoditized I think we are going to see an explosion of some of the most mind blowing art.
- flmontpetit 1y agoIt's already difficult enough to make a successful book adaptation, even WITH authorial intent. Can't imagine that hours of patchwork AI-generated video, with all its artifacting and consistency errors, will fare any better than "The Rings of Power".
- marcyb5st 1y agoI think not yet, but it is coming. I can see it using some form of PEFT so that the output becomes consistent with both the setting and the characters and then it is about generating over and over each short segment until you are happy with the outcome. Then you stitch them together and if you don't like some part you can always try to re-generate them, change the prompt, ...
- flmontpetit 1y agoI don't believe we will live to see the day where these models can replace a competent production team. At best they'll be what LLMs are to creative writing, which has so far only conclusively replaced low effort blogspam and fraud/plagiarism.
- impalallama 1y agoWell this is terrifying
- clarkcharlie03 1y agoGoogle's been coooooking
- aaroninsf 1y agoFunny but also illustrative issue: in the owl/badger video, the owl should fly silently. This is an interesting non-trivial problem of generalization and world-knowledge etc., but also? There's something somewhat sad about that slipping through; it makes me think, *no one involve in the production of this video, its selection, it passing review... etc., seemed to realize that it is one of the characteristic things about owls that you don't hear their wings. We have owls on our hill right now and see them almost every day and regularly seem them fly. It's magic, especially in an urban environment.
- thangalin 1y agoThe silent flight of an owl (BBC): https://www.youtube.com/watch?v=-WigEGNnuTE https://www.youtube.com/watch?v=-WigEGNnuTE Longer version: https://www.youtube.com/watch?v=-3ZnrhPtER8 https://www.youtube.com/watch?v=-3ZnrhPtER8