41 ms·
Sora 2
Video: https://www.youtube.com/watch?v=gzneGhpXwjU https://www.youtube.com/watch?v=gzneGhpXwjU
System card: https://openai.com/index/sora-2-system-card/ https://openai.com/index/sora-2-system-card/
- rhetocj23 1y ago[flagged]
- dang 1y agoCould you please stop posting unsubstantive comments and flamebait? You've unfortunately been doing it repeatedly. It's not what this site is for, and destroys what it is for. If you wouldn't mind reviewing https://news.ycombinator.com/newsguidelines.html https://news.ycombinator.com/newsguidelines.html and taking the intended spirit of the site more to heart, we'd be grateful.
- deleted 1y ago[deleted]
- dvngnt_ 1y agoAfter using Wan with comfyui, im uninterested in closed platforms. they lack the amount of control even if the quality might be better.
- kveykva 1y agoThe example prompt "intense anime battle between a boy with a sword made of blue fire and an evil demon demon" is super clearly just replicating Blue Exorcist https://en.m.wikipedia.org/wiki/Blue_Exorcist https://en.m.wikipedia.org/wiki/Blue_Exorcist
- greyk47 1y agoone of the example prompts is literally: Prompt: in the style of a studio ghibli anime, a boy and his dog run up a grassy scenic mountain with gorgeous clouds, overlooking a village in the distant background
- kossTKR 1y agoWow that is dark, after Ghiblis staunch stance on AI. These companies and their shareholders really are complete scum in my eyes, just like AI in miltech. Not because the tech isn't super interesting but because they steal years of hard work and pain from actual artists with zero compensation - and then they brag about it in the most horrible way possible, with zero empathy. Then comes losing the little humanity left the mainstream culture, exactly as Miyzaki said, leading to a dead cold and even more unjust society.
- martin-t 1y agoPower creates more power, money creates more money. Communism is tossing the frog into boiling water (tens millions of dead), capitalism is boiling it slowly (poor people in first world countries might not afford a dentist but they're not starving yet). We need a system that rewards work - human time and competence. There are really only 2 resources in the world - natural resources and human time. Everything else is built on top of those. And the people providing their time should be rewarded, not those who are in positions of power which allow them to extract value while not providing anything in return.
- martin-t 1y ago56 minutes, 4 downvotes, HN is truly full of temporarily embarrassed millionaires. Does anybody here really think rich people deserve to just get richer faster than any working person can? Does anybody really believe that buying up homes and companies and raking in money for doing absolutely nothing is what we should be rewarding? Then put your name behind it.
- astrange 1y agoHomes are depreciating assets. You can't get rich by "buying up homes and doing nothing" because you'd lose money. Nobody is doing this, although a bunch of confused people on social media believe BlackRock is doing it for some reason.
- aubanel 1y agoThat, and the dragon looking straight out of How to Train Your Dragon - I wonder if they have agreements with the right holders, or if they expect massive lawsuits to create free advertising for their launch.
- chris_wot 1y agoWell, look at Wikimedia. https://commons.wikimedia.org/wiki/File:This_Is_Fine_(meme).png https://commons.wikimedia.org/wiki/File:This_Is_Fine_(meme).... Here is a direct example of a derived work, to the point where the prompt is "n orange-brown anthropomorphic dog sitting in a chair at a table in a room that is engulfed in flames, happy dog sitting on chair at a table viewed from the side, dog with a hat, room is burning with fire all across the room". That's covered by Fair Use, I suppose they will argue this if they get sued. Interestingly, commons doesn't allow Fair Use, but the according to commons, "this is not a derived work". https://commons.wikimedia.org/wiki/Commons:Deletion_requests/File:This_Is_Fine_(meme).png https://commons.wikimedia.org/wiki/Commons:Deletion_requests...
- aubanel 1y agoThank you, interesting! I don't know that much about Fair use: if I understand well, the key is that the use should be "transformative", right? Am I correct in understanding that: - if the original "This is fine" meme was under copyright, the dog picture would be exempted from copyright by Fair use as it's a transformation - here it's not even needed since the original is not under copyright ("this is not a derived work")
- chris_wot 1y agoIt was a batshit insane decision, and a wrong one. Also: Commons doesn't allow for Fair Use images, so actually the decision was made that this wasn't transformative as it wasn't a derivative image. You tell me if that was a derivative image or not. I argued it was, and the argument was completely ignored.
- beernet 1y agoOverall, appears rather underwhelming. Long way to go still for video generation. Also, launching this as a social app seems like yet another desperate try to productize and monetize their tech, but this is the position big VC money forces you into.
- falcor84 1y agoI've perhaps been away from the scene for a bit, but I'm very impressed. To me this is absolutely "video generation", and I don't get your disdain for productization and monetization; last I checked this wasn't "Basic Research News".
- DetroitThrow 1y agoI don't think it's disdainful to point out the lack of PMF for a dedicated app for Sora, nor how its behind competitors who don't require a dedicated social app. No need to strawman the guy, I think it's okay to be reasonably critical of ideas still on this website. Inb4 make your own video model and see how easy it is
- msp26 1y agoThe voice quality in the generated vids is surprisingly awful.
- gmueckl 1y agoThat's the first thing I noticed, too. The first words you hear in the trailer sounds like someone ran the voice through a comb filter. It's so bad it made my skin crawl immediately.
- minimaxir 1y agoOpenAI apparently assumes that the primary users of Sora 2/the Sora app will be Gen Z, especially with the demo examples shown in the livestream. If they are trying to pull users from TikTok with this, it won't work: there's some nuance to Gen Z interests than being quirky and random, and if they did indeed pull users from TikTok then ByteDance could easily include their own image/video generators. Sora 2 itself as a video model doesn't seem better than Veo 3/Kling 2.5/Wan 2.2, and the primary touted feature of having a consistent character can be sufficiently emulated in those models with an input image.
- usaar333 1y agoPhysics seems better than veo 3 at least from demo videos
- bflesch 1y agoGood point. I think OpenAI lacks the cultural understanding that tiktok is providing their users not only with entertainment but also social things like trends, reviews, gossip, self-expression. These aspects are not included in the sora experience.
- rhetocj23 1y agoThis is going to sound crass but idc - OAI is just full of geeks, when what is needed is people who are more akin to hippies - thats pretty much what Apple was in the early days. Its no use building technology when its not married with the humanities and liberal arts.
- bflesch 1y agoIMO you're making a valid point, because there seems to be a disconnect between AI and tangible human benefits. The ChatGPT-as-therapy train has been nerfed after the bad publicity, and it is force-fed to people at their workplaces through copilot. I assume if you ask normal people how AI affects their lifes they'd think about annoying callcenter menus, deep fake porn and propaganda videos, and getting homework done. Not sure if any of this is a positive experience for the mind. It's 2025 and most speech controls for car navigation don't work, Siri is a pile of sh*t and millionaires are trying to convince us that we should either use their AI or a google which has significantly reduced the quality of their search result pages. It's like a false choice dilemma which allows back-to-the-roots companies such as Kagi to emerge, and I'm happy about it.
- deleted 1y ago[deleted]
- mdrzn 1y agoIf this is anything near the demo they have been released, this seems incredibly good at physics. Wow. Can't wait to try the new app.
- jsheard 1y agoSora 1 was also lauded as being incredibly good at physics based on the early cherry-picked examples. The phrase "world simulator" was thrown around a lot. That didn't last long once people finally got their hands on it though.
- DetroitThrow 1y agoThe space dog and ice skater demo make it seem still very close to Sora 1
- benjiro 1y agoKind of wondersome if they will start to combine LLM generation with actual world models/GPU engines. Imagine that your model generates the wireframes, the Engine generates the physics and then another model fills in the actual visuals, and gaps... So you have realistic physics and gaps are filled in... Will also help with image retention more, if objects moved behind each other.
- xenobeb 1y agoIt was so much more hyped than that. They made it sound like Hollywood was in big trouble. It is going to have the same problems as Midjourney. You just don't have that much control of the scene. The process is to make thousands of random variations and cherry pick the good stuff because you can't do anything else.
- techpression 1y agoThe demo on their homepage shows really bad physics. There’s a lot of it, but that doesn’t mean it’s correct. The hair of Sam looks like a paper cutout in almost every shot.
- spaceman_2020 1y ago
- fariszr 1y agoDid they make human voices sound robotic on purpose? Is that some kind of Ai fingerprinting? It's way too obvious
- minimaxir 1y agoIt's very hard for simultaneous good audio generation with video generation (simultaneous generation is necessary to maintain lip sync). Veo 3 et al also have flat monochannel audio, but not as bad as these Sora 2 demos.
- causal 1y agoIDK if the site is being hugged to death but I can only load the first video. Even in just one viewing there were noticeable artifacts, so my impression is that Veo is still in the lead here.
- qafy 1y agoYeah I am curious what the actual resolution of these videos will be. The launch videos on this link will only play in like 360p for me.
- S0und 1y agoI find it comical that OpenAI with all the power of CharGPT even them are unable to release an app for both iOS and Android at the same time. Wow, good marketing for Codex.
- aizk 1y agoThat is more of a statement of the complete dominance of iPhones among gen z.
- bigyabai 1y agoOr Sama's documented reverence for Apple products. We are talking about the guy who sold Tim Cook his AI for $0.00, he's not exactly got the horse drawing the cart here.
- drexlspivey 1y agoGoogle sold Tim Cook their search engine for $-25B per year
- gmuslera 1y agoNot even for all regions for iOS
- rd 1y agohttps://apps.apple.com/us/app/sora-by-openai/id6744034028 https://apps.apple.com/us/app/sora-by-openai/id6744034028 App link edit: CBN80W for an invite code
- throwup238 1y agoI downloaded the app but I get a "Sora is invite only" screen after logging in to my OpenAI account and asking for an invite code.
- Tiberium 1y ago> You can sign up in-app for a push notification when access opens for your account. You need to be in the US/Canada and wait for this notification, and when you get an invite you can start using it in the app and on sora.com. And apparently you get 4 more invite codes that you can share with anyone, e.g. Android users: > Android users will be able to access Sora 2 via http://sora.com http://sora.com once you have an invite code from someone who already has access
- qingcharles 1y agoIt's wild that I have a paid account but I have to scour the Internet to find someone else with a paid account and beg them for an invite code to use the product I already paid for. Make it make sense.
- gretch 1y agoOne thing that would make sense is for you to not pay any more. But if you do, that signals to the company this is all perfectly okay.
- Y_Y 1y agoDo you really want a "social" app for a firehose of high-fidelity slop?
- solfox 1y ago
- DetroitThrow 1y agoJust seeing the examples that I assumed are cherry picked, it seems like they're still behind on Google when it comes to video generation, the physics and stylized versions of these shots seem not great. Veo3 was such a huge leap and is still ahead of many of the other large AI labs.
- rushingcreek 1y agoThe most interesting thing by far is the ability to include video clips of people and products as a part of the prompt and then create a realistic video with that metadata. On the technical side, I'm guessing they've just trained the model to conditionally generate videos based on predetermined characters -- it's likely more of a data innovation than anything architectural. However, as a user, the feature is very cool and will likely make Sora 2 very useful commercially. However, I still don't see how OpenAI beats Google in video generation. As this was likely a data innovation, Google can replicate and improve this with their ownership of YouTube. I'd be surprised if they didn't already have something like this internally.
- deleted 1y ago[deleted]
- visarga 1y ago> the ability to include video clips of people and products as a part of the prompt and then create a realistic video with that This is something I would not like to see, I prefer product videos to be real, I am taking a risk with my money. If the product has hallucinated or unrealistic depiction it would be a kind of fraud.
- BeetleB 1y agoI believe existing laws already cover that issue.
- mepiethree 1y agoDeepfakes require zero work now
- pton_xd 1y agoSomeone remind me the benefits of mass produced fake videos again?
- ToucanLoucan 1y ago- Political propaganda - Scamming people at scale - Nonconsensual pornography - Juicing engagement metrics for fading social media sites - The ongoing destruction of truth as a concept in our increasingly atomized and divided world
- jablongo 1y agoI think the last one takes the cake.
- chis 1y agoI imagine it's incredibly useful for prototyping movies, tv, commercials before going to the final version. CGI will probably get way cheaper too with some hybrid approach. Obviously this will get used for a lot of evil or bad as well
- greyk47 1y agocan you imagine a billion dollar company promoting their new pre-vis app?
- jsheard 1y agoI feel like that's missing the point of pre-vis anyway, its purpose is to lay down key details with precision but without regard for fidelity (e.g. https://youtu.be/KMMeHPGV5VE https://youtu.be/KMMeHPGV5VE), a system with high fidelity but very loose control is the exact opposite of what they want.
- minimaxir 1y agoIt's fun: maybe not for everyone, but there's clearly sufficient interest in it. Whether said fun is "worth" the social and economic costs is a separate issue.
- 2OEH8eoCRo0 1y agoCan it generate an analog clock displaying a given time?
- martypitt 1y agoEven if it can't, that wouldn't make this demo any less impressive.
- gvv 1y agoAny idea if or when it will be available in EU? https://apps.apple.com/us/app/sora-by-openai/id6744034028 https://apps.apple.com/us/app/sora-by-openai/id6744034028 edit: as per usual it's not yet...
- aaroninsf 1y agoSomeone who doesn't follow the moving edge would be forgiven for being confused by the dismissive criticism dominating this thread so far. It's not that I disagree with the criticism; it's rather that when you live on the moving edge it's easy to lose track of the fact that things like this are miraculous and I know not a single person who thought we would get results "even" like this, this quickly. This is a forum frequented by people making a living on the edge—get it. But still, remember to enjoy a little that you are living in a time of miracles. I hope we have leave to enjoy that.
- cubefox 1y agoYeah. Just a few years ago, people here would have said stuff like that was decades away at best and pure science fiction at worst.
- Jordan-117 1y agoI remember having to tell people "Don't describe too many individual objects in your DALL-E 2 prompt, because the model has trouble separating discrete concepts and the features tend to blend together." Now we have photorealistic video with sound and, oh yeah, the model can generate an entire script and mini-plot on its own based on the most basic prompt.
- qoez 1y agoI know the comments here are gonna be negative but I just find this so sick and awesome. Feels like it's finally close to the potential we knew was possible a few years ago. Feels like a pixar moment when CG tech showed a new realm of what was possible with toy story
- m3kw9 1y agoNo doubt they can create Hollywood quality clips if the tools are good enough to keep objects consistent, example, coming back to the same scene with same decor and also emotional consistency in actors
- gretch 1y ago> keep objects consistent I think this is not nearly as important as most people think it is. In hollywood movies, everyone already knows about "continuity errors" - like when the water level of a glass goes up over time due to shots being spliced together. Sometimes shots with continuity errors are explicitly chosen by the editor because it had the most emotional resonance for the scene. These types of things rarely affect our human subjective enjoyment of a video. In terms of physics errors - current human CGI has physics errors. People just accept it and move on. We know that superman can't lift an airplane because all of that weight on a single point of the fuselage doesn't hold, but like whatever.
- ileonichwiesz 1y agoWater level in a glass changing between shots is one thing, the protagonist’s face and clothes changing is another.
- bbor 1y agoWell put. Honestly the actor part is mostly solved by now, the tricky part is depicting any kind of believable, persistent space across different shots. Based off of amateur outputs from places like https://www.reddit.com/r/aivideo/ https://www.reddit.com/r/aivideo/, at least! This release is clearly capable of generating mind-blowingly realistic short clips, but I don't see any evidence that longer, multi-shot videos can be automated yet. With a professional's time and existing editing techniques, however...
- modeless 1y agoI can see it being interesting to create wacky fake videos of your friends for a week or two, but why would people still be using this next year? I watch videos for two reasons. To see real things, or to consume interesting stories. These videos are not real, and the storytelling is still very limited.
- derac 1y agoI'm no Nostradamus, but I predict these models will be much better in a year.
- pr337h4m 1y agosoft porn
- qingcharles 1y agoYou only watch real things? Have you never watched a movie?
- modeless 1y ago> or to consume interesting stories
- FergusArgyll 1y agoIn the right hands it's a new art medium. Some (few, maybe) midjourney generations are serious art. So, for the same reason you'd go to a local art gallery
- bonoboTP 1y agoA lot of realslop is fake too. As in staged but pretended as real for rage bait or annoyance bait. Or stupid shaggy dog story videos, where it seems like the thing will happen any moment now and then nothing happens. One recent disillusionment for me was that lots of police body cam content is fake, as in basically amateur actors trying to enact a realistic police stop, they even put the usual bodycam numbers and letters and axos logo in the corner etc. And so many other videos of things happening in the street are more or less obviously fake and staged. Still 90% probably don't notice.
- jsnell 1y agoDoing this as a social app somehow feels really gross, and I can't quite put to words why. Like, it should be preferable to keep all the slop in the same trough. But it's like they can't come up with even one legitimate use case, and so the best product they can build around the technology is to try to create an addictive loop of consuming nothing but auto-generated "empty-calories" content.
- pavon 1y agoI see it more as recognizing that it will take time to be good enough for other use cases so for this release they are targeting it as just something to have fun with. After seeing LLMs crammed into everything whether it makes sense or not, I can appreciate that.
- dweekly 1y agoSo a social network that's 100% your friends doing silly AI things? I feel like this is the ultimate extension of "it feels like my feed is just the artificial version of what's happening my friends and doesn't really tell me anything about how they're actually faring."
- al_borland 1y agoSocial media also tends to highlight the best parts of people’s lives, creating unrealistic expectations and views for those consuming it and looking at their real life. Now social media won’t even be a highlight reel, but completely fabricated. I have to imagine there will be a rebellion against all of this at some point, when people simply can’t take the false realities anymore. What is the alternative? Ready Player One? The Matrix? Wall-E?
- m3kw9 1y agoWhich seem to level the play field, at least virtually
- al_borland 1y agoMaybe inside of a social network specially for AI, but a concerning number of people don't realize images and videos are AI, even when it's bad AI. As it gets better, and starts integrating the poster's image (like Sora 2), that's going to get even worse.
- kjs3 1y agoSome people use filters/photoshop to artificially juice their images; now they can use AI to artificially juice every aspect of their on-line presence.
- doctorhandshake 1y agoI built an MVP of this [1] with images (not video) and in more of an Instagram style (not tiktok) back in ‘22, with the tagline ‘What if social media were literally fake?’ I am bullish on this, albeit with major concerns in many domains. It was fun and addictive as hell with images. With video it will be wild. [1] https://hardwork.party/cheese/ https://hardwork.party/cheese/
- jablongo 1y agoSam Altman has made (for me) encouraging statements in the past about short-form video like TikTok being the best current example of misaligned AI. While this release references policies to combat "Doomscrolling and RL-sloptimization", it's curious that OpenAI would devote resources to building a social app based on AI generated short form video, which seems to be a core problem in our world. IMO you can't tweak the TikTok/YouTube shorts format and make it a societal good all of a sudden, especially with exclusively AI content. This is a disturbing development for Altman's leadership, and sort of explains what happened in 2023 when they tried to remove him... -> says one thing, does the opposite.
- bigyabai 1y agoSam Altman is a businessman. His job is to say whatever assuages his market, and that includes gaslighting you when you're disgusted by AI. If you never expected Altman to be the figurehead of principled philosophy, none of this should surprise you. Of course the startup alumni guy is going to project maligned expectations in the hopes of being a multi-trillion dollar company. The shareholders love that shit, Altman is applying the same lessons he learned at Worldcoin to a more successful business. There was never any question why Altman was removed, in my mind. OpenAI outgrew it's need for grifters, but the grifter hadn't yet outgrown his need for OpenAI.
- jablongo 1y agoTo be clear I'm not disgusted by AI in general, I'm disgusted by short form video and AI/ML in service of dopamine reward loop hacking.
- estearum 1y ago> His job is to say whatever assuages his market I understand the cynicism but this is in fact not the job of a businessman. We shouldn't perpetuate the pathological meme that it is.
- bnop 1y agoSo the job of a businessman is not to increase shareholder value?
- taytus 1y agoHonest question: What problem does this solve?
- lawlessone 1y agoLiberation of employers from the shackles of their employees
- martypitt 1y agoThing that was previously very expensive, manual and took a long time to do, and is done A LOT, is now made faster and cheaper by computers. Pretty much the same problem we all work on every day in $DAY_JOB.
- nextworddev 1y agoFun and games until someone uses a tool like this to scam your family
- deleted 1y ago[deleted]
- lawlessone 1y agoIt's ok, they're making the market for anti-ai tools much much bigger. (whether those tools work or not is a different issue)
- smith7018 1y agoOpenAI needing something to show to investors to say "See, this is why we need $1T."
- andybak 1y agoWhat problem does what solve? Video generation models in general or Sora 2 specifically?
- squidsoup 1y agoIt facilitates the generation of political propaganda.
- mempko 1y agoIt's obvious there is no way OpenAI can keep videos generated by this within their ecosystem. Everything will be fake, nothing real. We are going to have to change the way we interact with video. While it's obviously possible to fake videos today, it takes work by the creator and takes skill. Now it will take no skill so the obvious consequence of this is we can't believe anything we see. The worst part is we are already seeing bad actors saying 'I didn't say that' or 'I didn't do that, it was a deep fake'. Now you will be able to say anything in real life and use AI for plausible deniability.
- mmmrtl 1y agoI think that's the point... Then world coin comes to the rescue
- roxolotl 1y agoWorld coin is so delightfully dystopian. You could drop it wholesale into a superhero movie and it would be believable as the supervillain’s plot.
- kjs3 1y agoWe are going to have to change the way we interact with video. I doubt it will be for the better. The ubiquity of AI deepfakes just reenforces entrenchment around "If the message reinforces my preconceived notion, I believe it and think anyone who calls it fake is stupid/my enemy/pushing an agenda. If the message contradicts my preconceived notion, it's obviously fake and anyone who believes it is stupid/my enemy/pushing an agenda.". People don't even take the time to think "is this even plausible", much less do the intellectual work to verify.
- armchairhacker 1y agoRecord things with 2 cameras. Today's Sora can produce something that resembles reality from a distance, but if you look closely, especially if there's another perspective or the scene is atypical, the flaws are obvious. Perhaps tomorrow's Sora will overcome the the "final 10%" and maintain undetectable consistency of objects in 2 perspectives. But that would require a spatial awareness and consistency that models still have a lot of trouble with.
- whimsicalism 1y agoFind this sort of innovation far less interesting or exciting than the text & speech work, but it seems to be a primary driver of adoption for the median person in a way that text capability simply is not.
- liuliu 1y agoVideo generation is extremely exciting a.k.a. https://video-zero-shot.github.io/ https://video-zero-shot.github.io/ However, personalization (teleporting yourself into a video scene) is boring to me. At its core, it doesn't generate new experience to me. My experience is not defined by photos / videos I took on a trip.
- currymj 1y agoI also can't think of a reason why I would ever want to look at an AI generated video. however as they hint at a little in the announcement, if video generation becomes good enough at simulating physics and environments realistically, that's very interesting for robotics.
- mempko 1y agoIt's obvious there is no way OpenAI can keep videos generated by this within their ecosystem. Everything will be fake, nothing real. We are going to have to change the way we interact with video. While it's obviously possible to fake videos today, it takes work by the creator and takes skill. Now it will take no skill so the obvious consequence of this is we can't believe anything we see. The worst part is we are already seeing bad actors saying 'I didn't say that' or 'I didn't do that, it was a deep fake'. Now you will be able to say anything in real life and use AI for plausible deniability. I predict a re-resurgence in life performances. Live music and live theater. People are going to get tired of video content when everything is fake.
- minimaxir 1y agoThe Sora 2 livestream indicates that videos exported from the app will have visual watermarks.
- ileonichwiesz 1y agoSure, then you just pump it through another model that removes watermarks.
- saltyoldman 1y ago[flagged]
- jablongo 1y agoits well underway already
- password54321 1y agoalready happening: https://x.com/ken_wheeler/status/1954343731579994593 https://x.com/ken_wheeler/status/1954343731579994593
- mempko 1y agoI predict a re-resurgence in life performances. Live music and live theater. People are going to get tired of video content when everything is fake.
- nextworddev 1y agoOne would think, but people are spending less on live events due to costs
- volkk 1y agolikely because we haven't yet reached peak slop/exhaustion by slop. Soon enough...soon enough
- rvz 1y agoBuying lots of calls on Live Nation.
- nextworddev 1y agoMost of human crafted shorts / reels are already slop.
- simonw 1y agoAnyone with access able to confirm if you can start this with a still image and a prompt? The recent Google Veo 3 paper "Video models are zero-shot learners and reasoners" made a fascinating argument for video generation models as multi-purpose computer vision tools in the same way that LLMs are multi-purpose NLP tools. https://video-zero-shot.github.io/ https://video-zero-shot.github.io/ It includes a bunch of interesting prompting examples in the appendix, it would be interesting to see how those work against Sora 2. I wrote some notes on that paper here: https://simonwillison.net/2025/Sep/27/video-models-are-zero-shot-learners-and-reasoners/ https://simonwillison.net/2025/Sep/27/video-models-are-zero-...
- andrewguenther 1y agoYes, you can start with a still and a prompt
- andybak 1y agoI've got used to immediately checking availability. In this case - iPhone app is US + Canada only and the website is invite only. Going back to sleep. Wake me up when it's available to me.
- outlore 1y agoin a computer graphics course i took, we looked through how popular film stories were tied to the technical achievements of that era. for example, toy story was an story born from the new found ability to render plastics effectively. similarly, the sora video seems to showcase a particular set of slow moving scenes (or when fast, disappearing into fluid water and clouds) which seem characteristic of this technology at the current moment in time
- ChrisArchitect 1y agoMore discussion: https://news.ycombinator.com/item?id=45428122 https://news.ycombinator.com/item?id=45428122
- dang 1y agoComments moved thither. Thanks! Edit: looks like this post was actually first, so maybe we'll reverse the merge
- gorgoiler 1y agoImpressively high level of continuity. The only errors I could really call out are: 1/ 0m23s: The moon polo players begin with the red coat rider putting on a pair of gloves, but they are not wearing gloves in the left-vs-right charge-down. 2/ 1m05s: The dragon flies up the coast with the cliffs on one side, but then the close-up has the direction of flight reversed. Also, the person speaking seemingly has their back to the direction of flight. (And a stripy instead of plain shirt and a harness that wasn’t visible before.) 3/ 1m45s: The ducks aren't taking the right hand corner into the straightaway. They are heading into the wall. I do wonder what the workflow will be for fixing any more challenging continuity errors.
- fferen 1y agoVery first frame of the video: green digital text is messed up. Stopped watching after that :)
- yoavm 1y agoThe whole pool the ducks are racing at is a completely different pool when Sam starts talking.
- cogman10 1y agoThe snowmobiles were different in each cut. The shape, color, and style of the lights were different.
- fwip 1y agoNot sure if it counts as a continuity error, but in the example "Prompt: Martial artist doing a bo-staff kata waist-deep in a koi pond", his wooden staff changes shape several times, resembling a bow at points. That was the first example I noticed as "clearly AI."
- mNovak 1y agoThe Bo staff in the koi pond also seems to involve some impossible wrist movements
- tootie 1y ago
- willahmad 1y agoI wonder about the implications of this tech. State of the things with doom scrolling was already bad, add to it layoffs and replacing people with AI (just admit it, interns are struggling competing with Claude Code, Cursor and Codex) What's coming next? Bunch of people, with lots of free time watching non-sense AI generated content? I am genuinely curious, because I was and still excited about AI, until I saw how doom scrolling is getting worse
- m3kw9 1y agoI’m wondering how they really prevent uploads of other peoples faces if they take a clip of a video of another person. I’m sure Apple didn’t open up the 3d Face ID scanning to them to verify
- pixl97 1y ago>What's coming next? Bunch of people, with lots of free time watching non-sense AI generated content? Wasn't this always the outcome of the post labor economy? For this discussion lets just say that AI+Robots could replace most human labor and thinking. What do people do? Entertainment is going to be the number one time consumer.
- anshumankmr 1y agoPaid by what?
- quantumHazer 1y ago> just admit it, interns are struggling competing with Claude Code, Cursor and Codex They are not. This is false, zirp ended, this is the problem. Not LLMs.
- willahmad 1y agoOf course primary cause could be ZIRP, but AI definitely accelerated the problem. Interns at big tech maybe impacted less, because their systems are so complex, but when I look at job boards or talk with engineers I see they're mentioning interns less, AI assisted coding more. Bar for the interns is higher now, why do I need 3 interns to polish the product if I can complete 70% of the job with AI and hire 1 intern to fix other parts
- ElijahLynn 1y ago"download the Sora app" click takes me to the iPhone app store...
- m3kw9 1y agoI’m eagerly awaiting for some unexpected social problems this crops up
- sudohalt 1y agoNow videos will be generated on the fly based on your preference. You will never put your phone down, it will detect when your sad or happy and generate videos accordingly
- intended 1y agoThat dragon flew backwards at one point didnt it. Impressive that THAT was one of the issues to find, given where we were at the start of the year.
- adidoit 1y agoImpressive tech. Don't love the likely societal implications.
- joshdavham 1y agoWill something like Sora 2 actually be used in Hollywood productions? If so, what types of scenes? I imagine it won’t necessarily be used in long scenes with subtle body language, etc involved. But maybe it’ll be used in other types of scenes?
- gamegoblin 1y agoI saw a famous actor-director (can't remember who, but an A-list guy) said it would be super valuable even if you only use it for establishing shots. Like you have an exterior shot of a cabin, the surrounding environment, etc — all generated. Then you jump inside which can be shot on a traditional set in a studio. Getting that establishing shot in real life might cost $30K to find a location, get the crew there, etc. Huge boon to indie films on a budget, but being able to endlessly tweak the shot is valuable even for productions that could afford to do it IRL.
- esafak 1y agoProbably Ben Affleck. https://www.youtube.com/watch?v=ypURoMU3P3U https://www.youtube.com/watch?v=ypURoMU3P3U
- deelowe 1y agoWow. What an intelligent take. I would have never expected this from Ben Affleck. He seems extremely familiar with the technology and it's capabilities and limits.
- gamegoblin 1y agoSearched around and found it. It was actually Ashton Kutcher's interview with Eric Schmidt. Kutcher mentions the establishing shots, and I'd forgotten also points out the utility for relatively short stunt sequences. > Why would you go out and shoot an establishing shot of a house in a television show when you could just create the establishing shot for $100? To go out and shoot it would cost you thousands of dollars. > Action scenes of me jumping off of this building, you don’t have to have a stunt person go do it, you could just go do it [with AI].
- basisword 1y agoTens of billions in funding and they've just built a modern version of JibJab[1]. Can't wait to start receiving this in reply-all family emails. [1] https://youtu.be/z8Q-sRdV7SY?si=NjuyzL1zzq6IWPAe https://youtu.be/z8Q-sRdV7SY?si=NjuyzL1zzq6IWPAe
- simonw 1y agoThe main lesson I learned from the March ChatGPT image generation launch - which signed up 100 million new users in the first week - is that people love being able to generate images of their friends and family (and pets). I expect the "cameo" feature is an attempt at capturing that viral magic a second time.
- minimaxir 1y agoFortunately, you don't need permission from pets to use them in an AI video. (unless PETA objects)
- oulipo2 1y ago[flagged]
- minimaxir 1y agoThe Sora 2 system card claims Sora can resist generations of "political persuasion". https://cdn.openai.com/pdf/50d5973c-c4ff-4c2d-986f-c72b5d0ff069/sora_2_system_card.pdf https://cdn.openai.com/pdf/50d5973c-c4ff-4c2d-986f-c72b5d0ff...
- dragonwriter 1y ago> The Sora 2 system card claims Sora can resist generations of "political persuasion". To actually do that, it would need to have the evolving contextual knowledge of current events and the reasoning power to be able to identify prompts which, when requested, would likely have use in political persuasion, which would be a bigger breakthrough in AI than anything they are promoting. Conclusion: it can’t actually meaningfully do that, though it will probably reject some subset of prompts involving topics pre-identified to it as politically sensitive.
- oulipo2 1y ago"claims" <- that's the lie right there
- wilg 1y ago
- colonial 1y agoCool - now let's see how much it costs in compute to generate a single clip. (Also, notice how no individual scene is longer than a handful of seconds?)
- bergheim 1y agoWe are just heading for Lovely All TM. I kid. Art should require effort. And by that I mean effort on the part of the artist. Not environmental damage. I am SO tired of non tech friends SWOONING me with some song they made in 0.3 seconds. I tell them, sarcastically, that I am indeed very impressed with their endeavors. I know many people will disagree with me here, but I would be heart broken if it turned out someone like Nick Cave was AI generated. And of course this goes into a philosophical debate. What does it matter if it was generated by AI? And that's where we are heading. But for me I feel effort is required, where we are going means close to 0 effort required. Someone here said that just raises the bar for good movies. I say that mostly means we will get 1 billion movies. Most are "free" to produce and displaces the 0.0001% human made/good stuff. I dunno. Whoever had the PR machine on point got the blockbuster. Not weird, since the studio tried 300 000 000 of them at the same time. Who the fuck wants that? I feel like that ship in Wall-E. Let's invest in slurpies. Anyway; AI is here and all of that, we are all embracing it. Will be interesting to see how all this ends once the fallout lands. Sorry for a comment that feels all over the place; on the tram :)
- GuinansEyebrows 1y ago"if it's not worth [writing/playing/painting...], it's not worth [reading/listening/looking...]"
- bergheim 1y agoI had a friend over for my last birthday before going to a venue. He had a huge framed painting he had made. It made me cry. A prompt delivered by Amazon drones would obviously not be the same lovely moment. So yes, I agree.
- IncreasePosts 1y agoIt's fitting that they host the video on Youtube, since that is where all of their training data came from.
- stan_kirdey 1y agoThat could totally power next generation of green-screen techs. Generative actors may not find favorable response in the audiences; but SFX, decor, extras, environments that react to actors' actions - amazing potential.
- portaouflop 1y agoYou can already do really cool stuff in this area “old” tech like stable diffusion. Not realistic or anything but really cool looking/morphing images
- adventured 1y agoAt least in terms of realism, the image generation field is at the realism line now. Single frame generation with Wan 2.1 / 2.2 (and others) for example, will get you realism.
- zarzavat 1y agoI can see that future generations are going to think that I'm boomer for preferring the performances of real actors instead of AI slop. The music industry already went through this with AutoTune and we know how that turned out.
- poisonarena 1y ago>The music industry already went through this with AutoTune and we know how that turned out. they use it, everyone uses it, it got better to the point where most people dont know its used, ever heard of melodyne? well AI made it even better. And then there has been about 20 years of people using it even as their style of music, notably in hip hop, reggaeton, urbano, country, etc. Boomers like to think it was just an annoying fad in 2008-2011 or something, but it never went away, now everyone uses it, whether obvious or not
- r_lee 1y agoI don't get the autotune argument. It's like saying we shouldn't be using electronic instruments because it's not real or we shouldn't use digital audio instruments because they're not real etc. It's just a way to get different kind of sound. It won't make you good tracks.
- thebiglebrewski 1y agoCan this be used to make hyper-realistic video games, or it's not that real-time yet?
- dagaci 1y agoAmazing. iOS only, with region restrictions in 2025.
- asadm 1y agoconsidering legal foolishness of EU, this is the right move.
- TheAceOfHearts 1y ago> Sora is not available in Puerto Rico yet I love the casual reminds that we're second-class citizens each time a new technology gets released. Available in the US but always excluding Puerto Rico.
- GaggiX 1y agoThe model's quality is incredible, but more tools are needed to take advantage of its capabilities, this is kinda the magic of open models.
- barbarr 1y agoInstagram reels are gonna get crazy
- artursapek 1y agoYou see the one with the dolphin on the trampoline?
- ashu1461 1y agoThose `nature is amazing type of videos` are already flooded with AI
- MangoToupe 1y agoInteresting that they're going with a "copyright opt-out": https://www.reuters.com/technology/openais-new-sora-video-generator-require-copyright-holders-opt-out-wsj-reports-2025-09-29/ https://www.reuters.com/technology/openais-new-sora-video-ge... I guess copyright is pretty much dead now that the economy relies on violating it. Too bad those of us not invested into AI still won't be able to freely trade data as we please....
- alkonaut 1y agoHow far out are we from doing this in real time? What’s the processing/rendering time per frame?
- kachapopopow 1y agocould already do it in real time by dimming the lightbulbs of a city or two.
- neom 1y agohttps://deepmind.google/discover/blog/genie-3-a-new-frontier-for-world-models/ https://deepmind.google/discover/blog/genie-3-a-new-frontier...
- beders 1y agoCan I finally redo the Star Wars sequels with this? :)
- crims0n 1y agoDidn't Star Wars end in 2005?
- d--b 1y agoOk that's technically really impressive, and probably totally unusable in a real creativity context beyond stupid ads and politically-motivated deepfakes.
- deng 1y agoAs usual: impressive until you look close. Just freeze the frame and you see all the typical slop errors: pretty much any kind of writing is a garbled mess (look at the camera in the beginning). The horn of the unicorn sits on the bridle. The buttons on Sam's circus uniform hover in the air. There are candleholders with somehow candles inside as well as on top. The miniature instruments often make no sense. The conductor has 4 fingers on one hand and 5 on the other. The cheers of the audience is basically brown noise. Nedless to say, if you freeze the audience, hands are literally all over the place. Of course, everything conveniently has a ton of motion blur so you cannot see any detail. I know, I know. Most people don't care. How exciting.
- rendleflag 1y agoIs your complaint that it has errors? I mean look at what it can do. This is a freaking computer generating things from scratch based on a prompt. Two years ago, technology like this was so much worse and could only generate basic images and videos. Now it can generate visuals all from the text someone puts in. Anyone, literally anyone, can use it (eventually) to generate incredible scenes. Imagine the person who comes up with a short film about an epic battle between griffins and aliens...Or a simple story of a boy walking in the woods with their dog...Or a story of a first kiss. Previously people were limited to what they had at hand. They couldn't produce a video because it was too costly. Now they can craft a video to meet their vision. I do find it exciting.
- deng 1y ago> Is your complaint that it has errors? Well, yes? There's a reason why everything that was produced with these tools so far is garbage: because no one actually caring about their art would accept these things. Art is a deliberate thing, it takes effort. These tools are fine for company training videos and TikToks. Of course a few years ago this was science fiction. They are immensely impressive from a technical perspective. Two things can be true.
- bopbopbop7 1y agoThere is that magic word again, “eventually”. When is that? The same time we get warp drives?
- ascorbic 1y agoThis is super cool and fun and will almost certainly be really bad for society in loads of different ways. From the descriptions of all the guardrails they're needing to put in it seems like they know it too.
- bbor 1y agoGlad to see someone is looking out for a forest, here. A diverse host of excuses have cropped up to explain away the anxiety AGI brings, and I totally understand why. Yet again, today we stare into the abyss. Sora 2 represents significant progress towards [AGI]. In keeping with OpenAI’s mission, it is important that humanity benefits from these models as they are developed. This seems like a good time to remind ourselves of the original OpenAI charter: https://web.archive.org/web/20230714043611/https://openai.com/charter https://web.archive.org/web/20230714043611/https://openai.co... I wonder how exactly they reconcile the quote above with "We are concerned about late-stage AGI development becoming a competitive race without time for adequate safety precautions"...
- nurettin 1y agoI am not for or against AGI, but why is there anxiety around it? Do people simply hear sales rhetoric and assume that it can exist and will be used in order to dominate their lives?
- bbor 1y agoI'm not referencing sales rhetoric, I'm referencing scientific consensus. AGI will have the same kind of impact on our species as fire and electricity did. We stand at a crossroads between unimaginable success and enormous catastrophe...
- nurettin 1y agoWell, good luck with that, hopefully it will learn to spell blueberry.
- 1y ago
- haolez 1y agoOne use that occurred to me is that fans will be able to "fix" some movies that dropped the ball. For example, I saw a lot of people criticizing "Wish" (2023, Disney) for being a good movie in the first half, and totally dropping the ball in the last half. I haven't seen it yet, but I'm wondering if fans will be able to evolve the source material in the future to get the best possible version of it. Maybe we will even get a good closure for Lost (2004)! (I'm ignoring copyright aspects, of course, because those are too boring :D)
- BeetleB 1y agoOr just going to the Goofs section of a movie on IMDB, and fix the trivial issues (e.g. car had cracked window in earlier scene, and suddenly a normal window in another scene). Much more mundane, but useful!
- ronsor 1y ago> (I'm ignoring copyright aspects, of course, because those are too boring :D) You must understand that infinite copyright is the author's right, and AI companies must be sued for 50 trillion dollars.
- SkyBelow 1y agoMy issue is that the copyright aspect are what prevents me from using this as much as I otherwise would. About 6 months ago I asked a few different AIs if they could translate a song for me as a learning experience, meaning not a simple translation, but more a word by word explanation of what each word meant, how it was conjugated, any more musical/lyrical only uses that aren't common outside of songs, and so on. I was consistently refused on copyright grounds, despite this seeming a fair use given the educational nature. If I pasted a line of the lyrics at a time, it would work initially, but eventually I would need to start a new chat because the AI determined I translated too much at once. So in this one, if I wanted to ask it to create a video of the moment in Final Fantasy 6 when the bad guy wins, or a video of the main characters of Final Fantasy 7 and 8 having a sword duel, would it outright refuse for copyright reasons? It sounds like it would block me, which makes me lose a bit of interest in the technology. I could try to get around it, but at what point might that lead to my account being flagged as a trouble maker trying to bypass 'safety' features. I'm hoping in a few years the copyright fights on AI dies down and we get more fair use allowance instead of the tighter limitations to try to prevent calls for tighter regulation.
- qgin 1y agoVFX artists are definitely feeling the AGI / considering other career paths today.
- Banditoz 1y agoI genuinely don't understand the consistent rhetoric on this site of: > new AI feature/model comes out > "it's going to replace people in this field! they better start looking for a new job!!!" why is this a good thing?
- qgin 1y agoIt’s not a good thing, but it’s definitely a thing. Most of us here on HN are going to be affected by this.
- vultour 1y agoHow many more years do you think you'll need to keep saying this before it's actually true?
- qgin 1y agoNew grads are already having a tough time. My own expectation is that every recession or downturn from here on out, there will be the typical rounds of layoffs but without the typical increase in hiring afterwards. Maybe no “we replaced you with AI” moment, more of a “no new hiring” tendency.
- Retr0id 1y agoWho said it was a good thing?
- bopbopbop7 1y agoIs this AGI in the room with us now?
- myahio 1y agoNot with the way this thing renders hair (or any other high fidelity texture) https://x.com/GabrielPeterss4/status/1973090475486879818 https://x.com/GabrielPeterss4/status/1973090475486879818
- dragonwriter 1y ago“With Sora 2, we are jumping straight to what we think may be the GPT‑3.5 moment for video.” I think feeling like you need to use that in marketing copy is a pretty good clue in itself both that its not, and that you don’t believe it is so much as desperately wish it would be.
- echelon 1y agoThe Sora app squaring off against Meta's social video app is the real story here. Sora 2 itself looks and sounds a little poorer than Google Veo 3. (Which is itself not currently ranked as the top video model. The Chinese models are dominating.) I think Google, with their massive YouTube data set, is ultimately going to win this game. They have all the data and infrastructure in the world to build best-in-class video models, and they're just getting started. The social battle will be something completely different, though. And that's something that I think OpenAI stands a good chance at winning. Edit: Most companies that are confident of their image or video models stealthily launch it on the Model Arena a week ahead of the public model release. OpenAI did not arrange to do that for Sora 2. Nano Banana, Seedream/Seedance, Kling, and several other models have followed this pattern of "stealth ELO ranking, then reveal pole position". https://artificialanalysis.ai/text-to-video/arena?tab=leaderboard-image https://artificialanalysis.ai/text-to-video/arena?tab=leader... The fact that this model is about "friends" and "social" implies that this is an underpowered model. You probably saw a cherry picked highlight reel with a large VRAM context, but the actual consumer product will be engineered for efficiency. Built to sustain a high volume of cheap generations, not expensive high quality ones. A product built to face off against Meta. That model compete on the basis of putting you into videos with Pikachu, Mario, and Goku.
- CaptainOfCoit 1y ago> I think Google, with their massive YouTube data set, is ultimately going to win this game. I don't know, applying the same thinking to LLMs, Google should have been first and best with just text based LLMs too, considering the datasets they sit on (and researchers, among others the people who came up with attention). But OpenAI somehow beat them on that regardless.
- gainda 1y agoimpressive engineering that's hard to see as a net good for humanity. it doesn't spark optimism or joy about the future of engaging with the internet & content which was already at a low point. old is gold, even more so
- polishdude20 1y agoThere's something about the faces that looks completely off to me. I think it's the way the mouth and whole face moves when they talk.
- HarHarVeryFunny 1y agoYeah, the faces aren't right, and impressive as it is I'm getting icky "uncanny valley" vibes from this. CGI for fantasy stuff is unavoidable, but when it's stuff that could have been done by actors but is instead AI, then to me it just feels cheap and nasty - fake.
- bob1029 1y agoIt's the inaccuracy of things like shadows, sub-surface scattering and specular highlights. I think the shadow inaccuracy is what the human visual system is most sensitive to. These LLMs might make content that looks initially impressive but they are absolutely not performing physically based rendering or have any awareness of the lighting arrangement in these scenes. There are a lot of things they get right, but you only have to screw up one small element to throw the whole thing off. I am willing to bet that Unreal Engine 5 will continue to produce more realistic human faces than OAI ever can with these types of models. You cannot beat the effects of actually running raytracing in a PBR pipeline.
- dyauspitr 1y agoHow did they generate the videos with Sam Altman. Did they just provide a picture of his face and then use him in their prompts?
- rodonn 1y agoYou can use the "cameo" feature only with users who have gone through the cameo creation flow. Sama has an account and created a cameo likeness of himself. When you create your cameo you can choose who is allowed to make videos using it: "only me", "people I approve", "mutuals", or "everyone".
- kaicianflone 1y agoWhy is the video player so laggy?
- cubefox 1y agoRight? It constantly dropped frames for me (Firefox/Android).
- darkwater 1y agoLast famous words: > A lot of problems with other apps stem from the monetization model incentivizing decisions that are at odds with user wellbeing. Transparently, our only current plan is to eventually give users the option to pay some amount to generate an extra video if there’s too much demand relative to available compute. As the app evolves, we will openly communicate any changes in our approach here, while continuing to keep user wellbeing as our main goal.
- Workaccount2 1y agoSam will quickly learn that general users give -zero- thought to OpenAI well being. Nor be bothered that they should give it a thought.
- ambicapter 1y agoAI Sam Altman is terrifying, holy shit. Squarely in uncanny valley for me.
- benzible 1y agoCame here to say this myself. Would like to unsee that.
- neom 1y agoGoing to be an amazing source of training data, wait till they get it to real time and people are leaving their video camera open for AR features. OpenAI is about to have a lot of current real world image data, never mind the sentiment analysis.
- altcognito 1y agoI don't think they were limited for video training data. Gathering real world data is pretty easy, gathering curated information is a little more difficult.
- ishouldbework 1y ago[dead]
- bovermyer 1y ago"Thou shalt not create a machine in the likeness of a human mind."
- sciencejerk 1y agoAh, a holy scripture from the Orange Catholic Bible!
- saguntum 1y agoI wonder if they're going to license this to brands for heavily personalized advertisement. Imagine being able to see videos of yourself wearing clothes you're buying online before you actually place the order, instead of viewing them on a model. If they got the generation "live" enough, imagine walking past a mirror in a department store and seeing yourself in different clothes. Wild times.
- foota 1y agoThe latter would feel like actual scifi to me.
- larodi 1y agoits called Virtual Try On (VTO) and there are plenty of models going there for static gfx, it is very reasonable to expect soon emerge those for video VTO.
- shubb 1y agoAccurate virtual try on however is quite difficult, and users will quickly learn to distrust platforms that just generate something that"looks right". You can prompt with a normal size 8 dress and "kim jungle un wearing a dress" and it will show you something that doesn't help you understand whether that dress would fit or not. You can ask for a tube dress and it will usually give him a big bust to hold it up. It's not useful for the purpose of visualing fit. It will definitely be used for such just like image models already are for cheap tenu clothes, and our onions shopping experience will get worse. Maybe this needs purpose built models like vibe-net or maybe you cab train a general purpose model to do it, but if they were spending the effort necessary to do so they'd be calling it out.
- cyrialize 1y agoI'm fairly certain there is a scene in Minority Report just like this! Or at least, the advertisement says Tom Cruise's character's name. https://en.wikipedia.org/wiki/Minority_Report_(film) https://en.wikipedia.org/wiki/Minority_Report_(film)
- cindyllm 1y ago[dead]
- ath3nd 1y agoOpenAI is cooked. Absolutely cooked. After the disaster that was chatGPT4.001, study mode and now this: an impossibly expensive to maintain AI video slop copyright violater, their releases are uninspired and bland, and smelling of desperation. Making me giddy for their imminent collapse.
- dwa3592 1y agoI don't know if it's just me or other people are feeling it as well. I don't enjoy videos anymore (unless live sports). I don't enjoy reading on my monitor anymore, I have been going back to physical books more often. I am in my early thirties. The point is that sora2 demo videos seemed impressive but I just didn't feel any real excitement. I am not sure who this is really helping.
- marcofloriano 1y agoSame with me !
- greenavocado 1y agoPersonally I can't wait for super creative and novel indie film productions as film production will be more liberated from the grip of Hollywood and the influence of the upper classes in general. Especially once the Chinese make less-censored-to-Western-users models more available and even more so once people can run these things at home in some years.
- kobalsky 1y agothat sounds like clinical depresion, I'd check with my endocrinologist to get blood work done
- marcofloriano 1y agoEvery AI video demonstration is always about funny stuff and fancy situations. We never see videos on art, history, literature, poetry, religion (imagine building a video about the moment Jesus was born) ... ducks in a race !? Come on ... So much visual power, yet so little soul power. We are dying.
- fluoridation 1y agoWhat do you imagine a generated video about poetry would be? >Every AI video demonstration is always about funny stuff and fancy situations. The thing about AI slop is that by its very nature, unless it's heavily reined in by a human, it's invariably lowest common denominator garbage. It very likely will generate something you yourself could think of within the first five seconds of hearing the prompt, not some very clever take on it, so it can only work as a placeholder (AI as a replacement of stock images is great, for example) or to add background detail where it won't call attention to itself and its genericity. >imagine building a video about the moment Jesus was born Given there are multiple paintings on the subject, I very much doubt no one has generated something like that already.
- boh 1y agoThis is the kind of thing people get excited about for the first couple of months and then barely use it going forward. It's amazing how quickly the novelty of this amazing technology wears off. You realize how necessary meaning/identity/narrative is to media and how empty it gets (regardless of the output) when those elements are missing.
- tptacek 1y agoIf I was on the OpenAI marketing team I maybe wouldn't have included the phrase "and letting your friends cast you in their [videos]". It's a little chilling.
- minimaxir 1y agoThe livestream showed an interesting UX with Facebook-style permissions that make it so you very explicitly have to opt into this feature: https://bsky.app/profile/minimaxir.bsky.social/post/3m22zg2hhfs22 https://bsky.app/profile/minimaxir.bsky.social/post/3m22zg2h... Even moreso than Facebook tags, the person being cast can cause the deletion of the source video at any time.
- drcongo 1y agoThe AI generated Sam Altman doesn't look even vaguely human.
- echelon 1y agoI'm a software engineer and hobbyist actor/director. My friends are in the film industry and are in IATSE and SAG-AFTRA. I've made photons-on-glass films for decades, and I frequently film stuff with my friends for festivals. I love this AI video technology. Here are some of the films my friends and I have been making with AI. These are not "prompted", but instead use a lot of hand animation, rotoscoping, and human voice acting in addition to AI assistance: https://www.youtube.com/watch?v=H4NFXGMuwpY https://www.youtube.com/watch?v=H4NFXGMuwpY https://www.youtube.com/watch?v=tAAiiKteM-U https://www.youtube.com/watch?v=tAAiiKteM-U https://www.youtube.com/watch?v=7x7IZkHiGD8 https://www.youtube.com/watch?v=7x7IZkHiGD8 https://www.youtube.com/watch?v=Tii9uF0nAx4 https://www.youtube.com/watch?v=Tii9uF0nAx4 Here are films from other industry folks. One of them writes for a TV show you probably watch: https://www.youtube.com/watch?v=FAQWRBCt_5E https://www.youtube.com/watch?v=FAQWRBCt_5E https://www.youtube.com/watch?v=t_SgA6ymPuc https://www.youtube.com/watch?v=t_SgA6ymPuc https://www.youtube.com/watch?v=OCZC6XmEmK0 https://www.youtube.com/watch?v=OCZC6XmEmK0 I see several incredibly good things happening with this tech: - More people being able to visually articulate themselves, including "lay" people who typically do not use editing software. - Creative talent at the bottom rungs being able to reach high with their ambition and pitch grand ideas. With enough effort, they don't even need studio capital anymore. (Think about the tens of thousands of students that go to film school that never get to direct their dream film. That was a lot of us!) - Smaller studios can start to compete with big studios. A ten person studio in France can now make a well-crafted animation that has more heart and soul than recent by-the-formula Pixar films. It's going to start looking like indie games. Silksong and Undertale and Stardew Valley, but for movies, shows, and shorts. Makoto Shinkai did this once by himself with "Voices of a Distant Star", but it hasn't been oft repeated. Now that is becoming possible. You can't just "prompt" this stuff. It takes work. (Each of the shorts above took days of effort - something you probably wouldn't know unless you're in the trenches trying to use the tech!) For people that know how to do a little VFX and editing, and that know the basic rules of storytelling, these tools are remarkable assets that compliment an existing skill set. But every shot, every location, every scene is still work. And you have to weave that all into a compelling story with good hooks and visuals. It's multi-layered and complex. Not unlike code. And another code analogy: think of these models like Claude Code for the creative. An exoskeleton, but not the core driving engineer or vision that draws it all together. You can't prompt a code base, and similarly, you can't prompt a movie. At least not anytime soon.
- sumeruchat 1y agoShameless plug but I am creating a startup in this space called cleanvideo.cc to tackle some of the issues that will come with fake news videos. https://cleanvideo.cc https://cleanvideo.cc
- robotsquidward 1y agoIt's insanely impressive. At the same time, all these videos all look terrible to me. Still get extreme uncanny valley and literally makes me sick to my stomach.
- spaceman_2020 1y agoThis stuff works really well when you make something that's exaggerated reality, as in either an animation or a MTV-style music video I can't find the link now, but I saw a continuous shot video of a grocery store from the perspective of a fly. It was shot in the 90s music video style and looked so damn good. Some of the stuff being done by these guys is also a whole lot of fun (slightly NSFW and political content), and it fits the music video theme: https://www.youtube.com/watch?v=V4zwIhS2iZk https://www.youtube.com/watch?v=V4zwIhS2iZk
- jrop 1y agoAgree - leaps and bounds beyond anything I would have dreamed possible a few years ago...but... IDK, if I'm honest, the sound was way off too, not just the visuals. The music sounded detuned slightly, and the crowd noise was "crackly" etc. etc. It had a low-fidelity "quality" to it. Personally, I feel mixed feelings. I'm impressed, but I'm not looking forward to the new "movies" that are going to litter YouTube et al generated from this.
- unsnap_biceps 1y agoThey seem like they're low FPS videos. I wonder if they're rendering 24 FPS and it's mismatching youtube's 30 FPS and causing the weird stuttering.
- unethical_ban 1y agoI just had a thought: (spoilers Expanse and Hyperion and Fire Upon the Deep) Multiple sci-fi-fantasy tales have been written about technology getting so out of control, either through its own doing or by abuse by a malevolent controller, that society must sever itself from that technology very intentionally and permanently. I think the idea of AGI and transhumanism is that moment for society. I think it's hard to put the genie back in the bottle because multiple adversarial powers are racing to be more powerful than the rest, but maybe the best thing for society would be if every tensor chip disintegrated the moment they came into existence. I don't see how society is better when everyone can run their own gooner simulation and share it with videos made of their high school classmates. Or how we'll benefit from being unable to trust any photo or video we see without trusting who sends it to you, and even then doubting its veracity. Not being able to hear your spouse's voice on the phone without checking the post-quantum digital signature of their transmission for authenticity. Society is heading to a less stable, less certain moment than any point in its history, and it is happening within our lifetime.
- sys32768 1y agoI welcome a world where gullible people begin to doubt everything they see.
- deleted 1y ago[deleted]
- SeanAnderson 1y agoSheeeeeeeeeeesh. That was so impressive. I had to go back to the start and confirm it said "Everything you're about to see is Sora 2" when I saw Sam do that intro. I thought there was a prologue that was native film before getting to the generated content.
- iLoveOncall 1y agoI'm sorry but that's a gross exageration. If any of this was real film then I'd start a gofundme page for OpenAI to get better video production equipment and team because that would be laughably bad. If anything, it looks a lot worse than a lot of AI-generated videos I've seen in the past, despite being a tech demo with carefully curated shots. Veo 3 just blows this out of the water for example.
- SeanAnderson 1y agoIt's not an exaggeration to me? I literally stopped the video and went back to the start and re-read. You're more than welcome to speak about your opinions and experiences, but I'm speaking about mine. I'm over here thinking, "It felt like just yesterday I was laughing at trippy, incoherent videos of Will Smith eating spaghetti." I love the progress we're making. I love the competition between big companies trying to make the most appealing product demos. I love not knowing what the tech world is going to look like in six months. I love not thinking, "Man. The Internet was a cool invention to have grown up in, but now all tech is mundane and extractive." Every time I see AI progress I'm filled with childlike wonder that I thought was gone for good. I don't know if this represent SOTA for video generation. I don't care. In that moment I found it impressive and was commenting specifically on the joy I experienced watching the video. I find it frustrating to have that joy met with such negativity.
- ryandrake 1y agoDon't worry. AI is going to be monetized and extractive in no time. Just like Social Media went from "fresh, fun and cool new tech" to "how did we let this horrible beast take hold of the world," AI will take the same path. In 10 years or sooner, when 99.99% of what you read, hear, and watch is AI slop, you're going to post "This used to be a cool invention!" if there's even a place left for humans to post by that time.
- VagabundoP 1y agoI hate this vacant technology tbh. Every video feels like distilled advert mindless slop. There's still something off about the movements, faces and eyes. Gollum features.
- mrcino 1y agoSo, this is the AI Slop generator for the AI SlipSlop that Altman has announced lately. Brave new internet, where humans are not needed for any "social" media anymore, AI will generate slop for bots without any human interaction in an endless cycle.
- carabiner 1y agoCEO of Loopt makes a cameo at 1:28 in the youtube vid.
- mclightning 1y agoIt is very underwhelming. It seems like a step backward. Scam altman should be replaced before he runs the company to bankruptcy.
- iLoveOncall 1y agoShow me a coherent video that lasts more than 5 seconds and was generated with the model and maybe I'll start to care.
- the_duke 1y agoI haven't seen comments regarding a big factor here: It seems like OpenAI is trying to turn Sora into a social network - TikTok but AI. The webapp is heavily geared towards consumption, with a feed as the entry point, liking and commenting for posts, and user profiles having a prominent role. The creation aspect seems about as important as on Instagram, TikTok etc - easily available, but not the primary focus. Generated videos are very short, with minimal controls. The only selectable option is picking between landscape and portrait mode. There is no mention or attempt to move towards long form videos, storylines, advanced editing/controls/etc, like others in this space (eg Google Flow). Seems like they want to turn this into AITok. Edit: regarding accurate physics ... check out these two videos below... To be fair, Veo fails miserably with those prompts also. https://sora.chatgpt.com/p/s_68dc32c7ddb081919e0f38d8e006163d https://sora.chatgpt.com/p/s_68dc32c7ddb081919e0f38d8e006163... https://sora.chatgpt.com/p/s_68dc3339c26881918e45f61d9312e955 https://sora.chatgpt.com/p/s_68dc3339c26881918e45f61d9312e95... Veo: https://veo-balldrop.wasmer.app/ballroll.mp4 https://veo-balldrop.wasmer.app/ballroll.mp4 https://veo-balldrop.wasmer.app/balldrop.mp4 https://veo-balldrop.wasmer.app/balldrop.mp4 Couldn't help but mock them a little, here is a bit of fun... the prompt adherence is pretty good, at least. NOTE: there are plenty of quite impressive videos being posted, and a lot of horrible ones also.
- ch4s3 1y agoThat seems like an awful use of technology like this. I would imagine they mean to use that for serving ads, but how do you even generate conversations with ai slop plus product placements? I could see it working sometimes but I doubt it scales.
- deleted 1y ago[deleted]
- micromacrofoot 1y ago> slop plus product placements social media was heading this way before AI
- Computer0 1y agoAre users of the $20 tier really going to have to deal with that obnoxious bouncing watermark I wonder? The previous watermark could be cropped, but I often didn't feel the need to as I use it for fun, but that would make me not want to show anyone.
- ionwake 1y agoI think HN is too political like this tech is clearly amazing and it’s great they shipped it there should be more props even if it’s a billion dollar company.
- sailingparrot 1y agoYes the tech is amazing. But tech is not everything, after 20 years of social media, its pretty clear to everyone that those things can have large long term impact both positive and negative for society, discussing the potential impacts of the tech is not being "political", its just being interested in the future.
- ionwake 1y agoIm not sure, what if society only learns through hardships?
- sailingparrot 1y agoWell certainly if you don’t want us to discuss the possible implications ahead, then yes we can only close our eyes and learn from hardship once it’s there, but then what do we learn ? To not close our eyes next time ? We could just do that like now.
- ionwake 1y agoI think it depends on the time and place. So in this context it doesnt matter too much, its a billion dollar company, but if it was a guy with a project he just spent all year on, publishing it on HN for the first time, I would expect people to focus less on teh politics and more on the achievement which is something I dont see too often, unless the work is stellar. Perhaps there is just a high bar on HN
- unfitted2545 1y agoThere's a great lyric from ELUCID I think about when people say stuff like this: > I don't have the privilege to think everything ain't political
- doikor 1y agoDoes this survive panning the camera away for 5 to 10 seconds and then back? Or basic conversation scene with the camera cutting between being located behind either speaker once every few seconds? Basically proper working persistence of the scene.
- bsenftner 1y agoDude, this generation of AI video models are just starting to have basic camera production terms understood, and then it is exactly like LLM generation: it's a pull of a slot machine arm; you might get what you want, but that's "winning" and the slot machine only gives out winners one in every 100 pulls. Every possible thing that could not be right happens. For example, I'm working with a walking and talking character at this time using multiple AI video models and systems. Generated clips any length longer than 8 seconds risk rapid quality loss, but sometimes you can get up to 12-19 seconds without the generation breaking down. That means one needs to simulate a multiple camera shoot on a stage, so you can cut around the character(s) and create a longer sequence. But now you need to have multiple views of the same location to place your character(s) into - and current AI models can't reliably give you a "different angled views" of an environment. We just got consistent different views of characters, and it'll be another period until environments can be generally examined from any view. BUT, that's if people realize this is not in the models yet, and so far people are so fascinated by the fantasy violence and sexual content they can make nobody realizes you cannot simply "look left and right" in any of these models and that even works with consistency or reliability. There are workarounds, like creating one's entire set and environments in 3D models, for use as the backgrounds and starting frames, but that's now 3D media production + AI, and none of the AI tools generate media that even has alpha channels, and a lot of similar incompatibilities like that.
- carrozo 1y agoSora 2: Sloppy Seconds
- CSMastermind 1y agoAnyone have an invite they want to share with me lol.
- apetresc 1y agoIf anyone is feeling generous with one of their four invite codes, I'd really appreciate it. I'm at adrian@apetre.sc.
- animanoir 1y ago[dead]
- TheAceOfHearts 1y agoReally impressive engineering work. The videos have gotten good enough that they can grab your attention and trigger a strong uncanny valley feeling. I think OpenAI is actually doing a great job at easing people into these new technologies. It's not such a huge leap in capabilities that it's shocking, and it helps people acclimate for what's coming. This version is still limited but you can tell that in another generation or two it's going to break through some major capabilities threshold. To give a comparison: in the LLM model space, the big capabilities threshold event for me came with the release of Gemini 2.5 Pro. The models before that were good in various ways, but that was the first model that felt truly magical. From a creative perspective, it would be ideal if you could first generate a fixed set of assets, locations, and objects, which are then combined and used to bring multiple scenes to life while providing stronger continuity guarantees.
- lm28469 1y ago"open ai is so nice because they spoon feed us little pieces of dog shit every few days to acclimate us to swallowing huge quantities of dog shit every single hours of your life in the near future, praise our benevolent god Sam Altman", and you should cheer for it!
- sealeck 1y ago> I think OpenAI is actually doing a great job at easing people into these new technologies. It's not such a huge leap in capabilities that it's shocking, and it helps people acclimate for what's coming. This version is still limited but you can tell that in another generation or two it's going to break through some major capabilities threshold. This is a truly _wild_ way to describe "this version isn't much better than the previous one". Would you say "Apple's latest iPhone is a pretty small marginal improvement over the previous one, but it's useful to help peopel to acclimate for what's coming".
- NoahZuniga 1y agoTTS is horrible compared to Google's veo 3
- neilv 1y ago> And we're introducing Cameo, giving you the power to step into any world or scene, and letting your friends cast you in theirs. How much are they (and providers of similar tools) going to be able to keep anyone from putting anyone else in a video, shown doing and saying whatever the tool user wants? Will some only protect politicians and celebrities? Will the less-famous/less-powerful of us be harassed, defamed, exploited, scammed, etc.?
- notatoad 1y agoit seems like this is basically youtube's ContentID, but for your face. as long as you upload your "cameo" aka facial scan to them, they can recognize and control the generation of videos with it. if you don't give them your face, then they can't/won't. "Consent-based likeness. Our goal is to place you in control of your likeness end-to-end with Sora. We have guardrails intended to ensure that your audio and image likeness are used with your consent, via cameos. Only you decide who can use your cameo, and you can revoke access at any time. We also take measures to block depictions of public figures (except those using the cameos feature, of course). Videos that include your cameo—including drafts created by other users—are always visible to you. This lets you easily review and delete (and, if needed, report) any videos featuring your cameo. We also apply extra safety guardrails to any video with a cameo, and you can even set preferences for how your cameo behaves—for example, requesting that it always wears a fedora."
- neilv 1y agoIf this company's guardrails end up sufficiently working well in practice (note phrases like "intended", "take measures", and "preferences...requested", on things they can't do 100%)... there will be weak links elsewhere, letting similar computation be performed without sufficiently effective guardrails against abuse? How do we prepare for this? Societal adjustment only (e.g., disbelieving defamatory video, accepting what pervs will do)? Establishing a common base of cultural expectations for conduct? Increasing deterrence for abusers?
- deleted 1y ago[deleted]
- rvz 1y ago12,000+ "AI startups" have been obliterated.
- bgwalter 1y agoWhat is the target market for this? The videos are not good enough for YouTube. They are unrealistic, nauseating and dorky. Already now any YouTube video that contains a hint of "AI" attracts hundreds of scathing comments. People do not want this. Let me guess, the ultimate market will be teenagers "creating" a Skibidi Toilet and cheap TikTok propaganda videos which promote Gazan ocean front properties.
- LarsDu88 1y agoI really hope they have more granular APIs around this. One use case I'm really excited about is simply making animated sprites and rotational transformations of artwork using these videogen models, but unlike with local open models, they never seem to expose things like depth estimation output heads, aspect ratio alteration, or other things that would actually make these useful tools beyond shortform content generation.
- jp57 1y agoPrediction: we'll see at least one Sora-generated commercial at the Super Bowl this year.
- OfflineSergio 1y agoWhile the quality of what I'm seeing is very nice for AI generated content (I still can't believe it) but the fact thay they are mostly showing short clips and not a long connected consistent video makes it less impressive.
- squidsoup 1y agoA little tangential to this announcement, but is anyone aware of any clean/ethical models for AI video or image generation (i.e. not trained on copyright work?) that are available publicly?
- egeres 1y agoI wonder how this will affect the large cinema production companies (Disney, WB, Universal, Sony, Paramount, 20th century...). The global film market share was estimated to be 100B in 2023. If the production cost of high FX movies like Avengers Infinity War goes down from 300M$ to just 10K$ in a couple of years, will companies like Disney restrain themselves to just release a few epic movies per year? Or will we be flooded with tons of slop? If this kind of AI content keeps getting better, how will movies sustain our attention and feel 'special'? Will people not care if an actor is AI or real?
- ashu1461 1y agoThis is a good comparison thread of capabilities of sora vs sora 2 https://x.com/mattshumer_/status/1973085321928515783 https://x.com/mattshumer_/status/1973085321928515783
- seydor 1y agoSince Agi is cancelled, at least we have shopping and endless video
- clgeoio 1y ago> Concerns about doomscrolling, addiction, isolation, and RL-sloptimized feeds are top of mind—here is what we are doing about it. > We are giving users the tools and optionality to be in control of what they see on the feed. Using OpenAI's existing large language models, we have developed a new class of recommender algorithms that can be instructed through natural language. We also have built-in mechanisms to periodically poll users on their wellbeing and proactively give them the option to adjust their feed. So, nothing? I can see this being generated and then reposted to TikTok, Meta, etc for likes and engagement.
- alberth 1y agoWhy do you have to download an app to use Sora 2 (vs it being available on the web like ChatGPT)?
- samuelfekete 1y agoThis is a step towards a constant stream of hyper-personalised AI generated content optimised for max dopamine.
- dwd 1y agoI hate to be right sometimes (got downvoted back in 2023) https://news.ycombinator.com/item?id=38705857 https://news.ycombinator.com/item?id=38705857 https://news.ycombinator.com/item?id=38706074 https://news.ycombinator.com/item?id=38706074
- taberiand 1y agoThe Torment Nexus is a Skinner box
- pawelduda 1y agoIt's far from sustainable (for now)
- kfarr 1y agoAssuming you have to generate new content for each viewer second watched yes it won't pencil out. But if you have a library of tons of content you can keep building out...
- ares623 1y agoKids will go to School V2 and have absolutely nothing in common to talk about because each one will have completely unique media entertainment at home.
- fersarr 1y agoOnly iphone...
- nycdatasci 1y agoWhat makes TikTok fun is seeing actual people do crazy stuff. Sora 2 could synthesize someone hitting five full-court shots in a row, but it wouldn’t be inspiring or engaging. How will this be different than music-generating AI like Suno, which doesn't have widespread adoption despite incredible capabilities?
- punkbit 1y agoIt's hard to believe, but some people enjoy. On the other hand, some popular content on TikTok is probably worse than AI generated content and that's another problem...
- cesarvarela 1y agoConsidering that much of the TikTok content you mention is staged or heavily edited, this skips the make-believe.
- dolebirchwood 1y agoThis makes me less excited about the future of video, not more. It's technically impressive, but all so very soulless. When everything fake feels real, will everything real feel fake?
- nalimtasseb 1y agoTruly wonder if there will be some kind of renaissance in the video making domain when all settles down and this becomes the new normal.
- rhetocj23 1y agoTastes and preferences are dynamic. It will certainly happen.
- bsenftner 1y agoThe ease of creating visually titillating media, coupled with the difficultly of consistency works against the creation of narrative media. I sure hope we don't get a generation of non-narrative beautiful slop.
- deleted 1y ago[deleted]
- amelius 1y agoNicely cherry-picked.
- ezomode 1y agofull-on productisation effort -> no AGI in sight
- mrcwinn 1y ago[flagged]
- Josh5 1y agoEveryone has the widest eyes in these Sora videos.
- yahoozoo 1y agoSam still pretending they’re close to AGI in the trailer lmao
- FullMetul 1y agoMaybe by Sora 3 they will have scene consistency. Gah it's so jarring to me that the poll the racing ducks are in just randomly changes. My brain can tell it's not consistent scene to scene and feels so jank.
- groos 1y agoWhat is the point? Who wants to watch these videos?
- Havoc 1y agoThat sure seems to be getting close to something usable for movies...kinda. Sam looks weirdly like Cillian Murphy in Oppenheimer in some shots. I wonder whether there was dataset bleedover from that.
- cogman10 1y agoI've seen a lot of "this is impressive" but I'm not really seeing it. This looks to suffer from all the same continuity problems other AI videos suffer from. What am I looking at that's super technically impressive here? The clips look nice, but from one cut to the next there's a lot of obvious differences (usually in the background, sometimes in the foreground).
- paulcole 1y agoAs a gauge for how seriously I should take your critique: How many hours a week are you actively using AI tools yourself? What percentage of public comments that you’ve made about AI tools have been skeptical or critical?
- cogman10 1y ago> How many hours a week are you actively using AI tools yourself? 2 or 3. Mostly LLMs to check code. > What percentage of public comments that you’ve made about AI tools have been skeptical or critical? Probably around 90%. So sell me. Why is this super impressive? I'm happy to admit that I'm pretty pessimistic about AI. I have an eye for continuity issues, they are pretty obvious to me. Am I just too focused on that sort of a thing?
- paulcole 1y ago[flagged]
- cogman10 1y agoI agree. It's really interesting that computers can make videos from sentences. This, however, isn't the first AI capable of doing that correct? Didn't Sora 1, Veo, and others come out before this making videos from sentences? Surely what makes this impressive isn't "This does what other people have done, including us Open AI". Since you edited, I will to respond to your inflammatory edit > Shocker lol > Like a guy who hates tomatoes, bread, cheese, and pepperoni going, “This pizza sucks.” This is a technology, don't reduce it down to preferences. There are obvious flaws with the video generating tech and one of the most annoying parts of talking to AI enthusiasts is the fact that they are unable to engage in honest dialog. All AI is amazing alien tech, it's always flawless and perfect. What I've seen with AI video in the past is that they can make impressive on first glance looking videos but when you dig into them things are "off". Further, the continuity only lasts for 1 continuous shot. That creates videos where every 5 seconds you see a new shot and in the same setting things tend to drastically change. Sora 2 appears to have all those problems, it doesn't appear to have solved any of them. That's why I ask "What's super impressive about this". The same way I'd ask "What's super impressive about ChatGPT 5 vs 4". Snarkly saying "What are you talking about, it is a fucking realistic chat that can write a short story!" doesn't convince or impress me or anyone else that's a skeptic. I'm not enthralled by AI. I'll happily use it when it makes sense and as it improves I'll probably use it more. For this, I don't see a significant improvement over the prior state of art.
- umrashrf 1y agohey @simoncion looks like they are doing this for self-promotion that's against the site's guidelines
- dcreater 1y agoMatrix here we come!
- Aeolun 1y agoClicking a link on the OpenAI dashboard and beeing greeted with a full page of scandily clad women was certainly not what I expected to see when opening Sora..
- aabhay 1y agoYou think too highly of us (humans)
- nopinsight 1y agoOpenAI launches Sora 2 in a consumer app to collect RL feedback en masse and improve their world models further. Their ultimate goal is physical AGI, although it wouldn’t hurt them if the social network takes off as well.
- btbuildem 1y agoThey're really playing loose with copyright: you have to actively opt out for them to not use your IP in the generated videos [1] Tangentially related: it's wild to me that people heading such consequential projects have so little life experience. It's all exuberance and shiny things, zero consideration of the impacts and consequences. First Meta with "Vibes", now this. 1: https://www.gurufocus.com/news/3124829/openai-plans-to-launch-sora-2-video-generator-amid-copyright-concerns https://www.gurufocus.com/news/3124829/openai-plans-to-launc...
- crazygringo 1y agoDo you have a better source for that? The footer to that article explicitly states the article is bot-generated.
- Barbing 1y agoLooks like WSJ broke the news: “OpenAI’s New Sora Video Generator to Require Copyright Holders to Opt Out” https://www.wsj.com/tech/ai/openais-new-sora-video-generator-to-require-copyright-holders-to-opt-out-071d8b2a https://www.wsj.com/tech/ai/openais-new-sora-video-generator... And Reuters covered their coverage minus the paywall: https://www.reuters.com/technology/openais-new-sora-video-generator-require-copyright-holders-opt-out-wsj-reports-2025-09-29/ https://www.reuters.com/technology/openais-new-sora-video-ge...
- ls612 1y agoI mean Grok has been free rein for copyrighted characters for over a year now and nobody’s sued them.
- Palmik 1y ago> people heading such consequential projects have so little life experience What do you mean by life experience here and how can you tell they have little of it?
- Lucasoato 1y ago> this app is not available in your country or region
- tonyabracadabra 1y agoIf Sora 2 is aiming for AI‑Tok, ScaryStories Live is the jump-scare cousin: real‑time POV horror from a photo + a sentence. No film school, no GPU farm—just “upload face, pick fear level, go.” It’s less cinema, more haunted mirror, and it ships in seconds. scarystories.live
- jug 1y agoI feel so bad for the climate now.
- bamboozled 1y agoSoon, you won't even have to do anything to post a video of yourself doing something "interesting" on social media, what at time to be alive. There would for sure be large swathes of people who would just lie about what they're doing and use AI to make it seem like they're skateboarding, or skiing or whatever at a pro or semi-pro level and have a lot of people watch it.
- mscbuck 1y agoI can't help but see these technologies and think of Jeff Goldblum in Jurassic Park. My boss sends me complete AI Workslop made with these tools and he goes "Look how wild this is! This is the future" or sends me a youtube video with less than a thousand views of a guy who created UGC with Telegram and point and click tools. I don't ever think he ever takes a beat, looks at the end product, and asks himself, "who is this for? Who even wants this?", and that's aside from the fact that I still think there are so many obvious tells with this content that make you know right away that it is AI.
- afavour 1y agoThis was my reaction when I saw Meta’s “Vibes” app. Who wants to browse a stream of exclusively AI generated videos? Obviously Meta wants that because it’s a lot cheaper than actually paying real people to make content… but it’s slop.
- bonoboTP 1y agoThis is not the final target. It's video generation now, but that's just a stepping stone. The real thing is that learning a generator is also learning a prior over videos, and hence over how the world works. The real application of this will be word models, vision-language action models, spatial AI and robotics. Basically a kind of learned simulator in which to plan and imagine possible futures, possible actions and affordances etc. Video models could become a spatial reasoning platform too. A recent paper by deepmind (using veo3) showed that video models can perform many high level vision tasks out of the box. Don't think it's going to end here at some slop feed.
- afavour 1y agoSure. But why do I, as a user, want to download Vibes today?
- gyomu 1y ago> This is not the final target The final target of these "world models" on a 20 year horizon is entirely unmanned factories taking over the economy, and swarm of drones and robots fighting wars and policing citizens. This is why hundreds of billions are poured into these things, cute Ghibli style videos and vacuum robots wouldn't be worth this much money otherwise.
- natiman1000 1y agoThe fact that no one talking about how it compares against Veo tells me everything I need to know. This page is now filled with some bots!
- mostMoralPoster 1y ago[dead]
- baby 1y agoNo android app right?
- minimaxir 1y agoThis Sora 2 generation of Cyberpunk 2077 gameplay managed to reproduce it extremely closely, which is baffling: https://x.com/elder_plinius/status/1973124528680345871 https://x.com/elder_plinius/status/1973124528680345871 > How the FUCK does Sora 2 have such a perfect memory of this Cyberpunk side mission that it knows the map location, biome/terrain, vehicle design, voices, and even the name of the gang you're fighting for, all without being prompted for any of those specifics?? > Sora basically got two details wrong, which is that the Basilisk tank doesn't have wheels (it hovers) and Panam is inside the tank rather than on the turret. I suppose there's a fair amount of video tutorials for this mission scattered around the internet, but still––it's a SIDE mission! Everyone already assumed that Sora was trained on YouTube, but "generate gameplay of Cyberpunk 2077 with the Basilisk Tank and Panam" would have generated incoherent slop in most other image/video models, not verbatim gameplay footage that is consistent. For reference, this is what you get when you give the same prompt to Veo 3 Fast (trained by the company that owns YouTube): https://x.com/minimaxir/status/1973192357559542169 https://x.com/minimaxir/status/1973192357559542169
- Klonoar 1y ago> Everyone already assumed that Sora was trained on YouTube Doesn't this already answer your question...? "Let's Play" type videos and streams have been a thing for years now, even for more obscure games. It very well could've been trained on Cyberpunk videos of that mission.
- minimaxir 1y agoIt's hard for me to believe that the model coherently memorized both the video and audio of a relatively obscure Let's Play, and that a simple prompt was enough to surface it (the use of the term "Basilisk tank" would also likely not be in video metadata either). That is the reason the person who made that tweet, who has far more prompting experience than myself, was shocked.
- Klonoar 1y agoIt’s hard for you to believe, sure, and I recognize the context of who tweeted it. I still maintain that’s the kernel it’s getting it from. It’s impressive, I’m just not really shocked by it as a concept.
- davidmurdoch 1y agoI just asked GPT 5 to generate an image of as person. I then asked it to charge the color of their shirt. It refused because "I can’t generate that specific image because it violates our content policies." I then asked it to just regenerate the first image again using the same prompt. It replied "I know this has been frustrating. You’ve been really clear about what you want, and it feels like I’m blocking you for no reason. What’s happening on my side is that the image tool I was using to make the pictures you liked has been disabled, so even if I write the prompt exactly the way you want, I can’t actually send it off to generate a new image right now." If I start a new chat it works. I'm a Plus subscriber and didn't hit rate limits. This video gen tool will probably be even more useless.
- Oarch 1y agoHaving AI explain policy violations in depth with the user could be a nice idea
- newZWhoDis 1y agoWe live in an absurd era where "AI Safety" means "AI that doesn't listen to the human telling it what to do". It'll all be rather funny in retrospect.
- cindyllm 1y ago[dead]
- danielscrubs 1y agoIt will be funny if it isn’t social engineering. But if we find it drifts further and further from the truth in cases of biases in news articles, image generation and others we will find ourselves bombarded with historical deviances where everyone can be nudged to anything. All in the name of safety.
- chii 1y agothat's why the AI capabilities should be as decentralized and "localized" as possible - aka, i want to own the hardware and software for LLM, image generation, etc etc. Until these ai capabilities are as neutral and un-discriminatory as electricity, centralized production means centralized control and policies. Imagine if you are not allowed to use your electricity to power some appliances, because the owner of the power-plant feels it's not conducive to their agenda.
- FrustratedMonky 1y agoYeah, we've "plateaued" all right.
- anshumankmr 1y agoI think someone had called it many months back (and in fact I felt it too) that the feed for Sora seemed very much like a social media app. Then the only thing left was to make it into vertical scrolling with videos and voila you have your tiktok clone.
- elpakal 1y agoWish I was cool enough to have an invite code. Oh well, as an iOS build nerd next best thing I can do is inspect their ipa I guess. Interesting that they have some pretty big duplicate mp4s nobody caught in NoFaceDesignSystemBundle: cameo_onboarding_0.mp4 & create_ifu_1.mp4 | 7.3MB and cameo_onboarding_2.mp4 & create_ifu_0.mp4 | 5.2MB. Also I find it neat that they still include an iOSMath bundle (in chatGPT too), makes me wonder how good their models really are at math.
- outside1234 1y agoThis is going to be a disaster. We are never going to be able to trust a video again and in short order propagandists are going to be using this to generate god knows what.
- _ZeD_ 1y agoSora 2: Frato
- wltr 1y agoFrom watching the video I have an impression that these guys just want to appear cool, and the product looks like that too. To appear to be very cool, for people who won’t ever use it, apparently. Same impression I’ve got from watching that promo with Jony Ive. Beautiful, and don’t you dare to think it through.
- LocalH 1y agoWe're cooked.
- baalimago 1y agoThey can't even be consistent within their own launch video. Consistency is by far the biggest issue with generative AI. How can a professional studio work with scenes which has continuity errors on every single shot? And if it's not targeting professionals, who is it for?
- ksynwa 1y agoThe common thread I am seeing with replacing creative work with AI is that of lowering the bar of acceptability and counterbalancing that with the (potential) savings from taking human labour out of the equation. The cost of labour is not just the raw cost but also the bargaining power that they can exercise by going on strikes etc. From my limited understanding, creatives seem to have more unions than programmers given that I have heard of at least two strikes from voice actors and writers and none from the tech sector. So it should be a win-lose for those who profit off of videos without taking part in the labour process of making one and lose-lose for everyone else.
- BoorishBears 1y ago> is that of lowering the bar of acceptability Yes. > counterbalancing that with the (potential) savings No. It's all about personalization. Even with all the money in the world you couldn't sit a filming crew, VFX specialist, foley artist, and voice actors next to every user of your app, ready to produce new content in 60 seconds. I don't get why this keeps being framed as a labor thing, it's unlocking genuinely new forms of interactive media.
- ksynwa 1y agoWhat kind of personalisations are you hoping to see with this tech? > I don't get why this keeps being framed as a labor thing It's inextricably linked with labour. That doesn't mean that labour is only factor but it's an important one nonetheless.
- BoorishBears 1y ago
- dang154f4g 1y ago[dead]
- tminima 1y agoI feel that this is a data collection activity (and thus, more advanced future models and usecases) disguised as a social media. People will provide feedback in the form of clicks/views on AI generated content (better version of RLHF) on unverified/subjective domains. Biggest problem OpenAI has is not having an immense data backbone like Meta/Google/MSFT has. I think this is step in that direction -- create a data moat which in turn will help them make better models.
- Gnarl 1y agoAmazing that even Sora2 can't make Sam Altman not look like a w@nker.
- jack_riminton 1y agoLets take a step back and realise how incredible this is (I'm sure there are plenty of other `ackshually` comments) Can it do Will Smith eating spaghetti? (I can't get access in UK)
- etrvic 1y agoIn light of some comments and videos here, I’d like to morbidly announce that I can no longer distinguish between AI videos and real ones. However, I’ll take this as an opportunity to move from short-form content to long-form, since it seems that space hasn’t yet been hijacked by AI.
- mavamaarten 1y agoUgh. While technically extremely impressive, I'm so tired of the slop. Every AI content generation tool should have a watermarking system in place, and sites like YouTube should have a way to filter out AI generated content from search results with the press of a button. Ever since the launch of Veo, there's already so much AI slop videos on YouTube that it becomes hard to find real videos sometimes. I'm tired, boss.
- taikahessu 1y agoEntering code 123456 reveals Sora 2 is only available in US/Canada region.
- TechSquidTV 1y agoNot related to Sora but, I have been looking for / hoping for an AI powered motion tracking solver. I've used Blender and Mocha in AE and both still require quite a bit of manual intervention, even in very simple scenes. I saw some promnise with the Segment Anything model but I haven't seen anyone yet turn it into a motion solver. In fact I'm not sure if can do that at all. It may be that we need to use an AI algorithm to translate the video into a more simple rendition (colored dots representing the original motion) that can then be tracked more traditionally.
- kkukshtel 1y agoYou should look at this Google paper (came out a few days ago): https://video-zero-shot.github.io/ https://video-zero-shot.github.io/
- Awesomedonut 1y agoTheir anime vid gen is really, really impressive. The results I've seen aren't /good/ from an industry-standard (nothing compared to the likes of the Demon Slayer movie I watched in theatres recently), but I legitimately couldn't tell that it was AI-generated. Massive step up from Sora 1 and other vid gen models. Here's to hoping that the industry will adapt to have it aid animators for in-betweening and other things that supplement production. Anime studios are infamously terrible with overworking their employees, so I legitimately see benefits coming from this tool if devs can get it to function as proper frame interpolation (where animators do the keyframes themselves and the model in-betweens).
- sandspar 1y agoSora 2 is a lot of fun. Using it feels like a glimpse into the future. The last time I felt this was with the Sesame voice demo.
- type0 1y agoit should be renamed into sore ai
- nickbettuzzi 1y agohi there! would love an invite code if anyone reading this has a spare. really interesting stuff— thank you in advance! email is nick@usmobile.com
- wantering 1y agoLatest Invite Codes Community-shared Sora invite codes updated in real-time https://sorainvitecode.org/ https://sorainvitecode.org/
- baby6343 1y ago[dead]