[HN Gopher] Sora 2
___________________________________________________________________
Sora 2
Video: https://www.youtube.com/watch?v=gzneGhpXwjU System card:
https://openai.com/index/sora-2-system-card/
Author : skilled
Score : 427 points
Date : 2025-09-30 16:55 UTC (6 hours ago)
(HTM) web link (openai.com)
(TXT) w3m dump (openai.com)
| dvngnt_ wrote:
| After using Wan with comfyui, im uninterested in closed
| platforms. they lack the amount of control even if the quality
| might be better.
| kveykva wrote:
| The example prompt "intense anime battle between a boy with a
| sword made of blue fire and an evil demon demon" is super clearly
| just replicating Blue Exorcist
| https://en.m.wikipedia.org/wiki/Blue_Exorcist
| greyk47 wrote:
| one of the example prompts is literally: Prompt: in the style
| of a studio ghibli anime, a boy and his dog run up a grassy
| scenic mountain with gorgeous clouds, overlooking a village in
| the distant background
| kossTKR wrote:
| Wow that is dark, after Ghiblis staunch stance on AI.
|
| These companies and their shareholders really are complete
| scum in my eyes, just like AI in miltech.
|
| Not because the tech isn't super interesting but because they
| steal years of hard work and pain from actual artists with
| zero compensation - and then they brag about it in the most
| horrible way possible, with zero empathy.
|
| Then comes losing the little humanity left the mainstream
| culture, exactly as Miyzaki said, leading to a dead cold and
| even more unjust society.
| martin-t wrote:
| Power creates more power, money creates more money.
|
| Communism is tossing the frog into boiling water (tens
| millions of dead), capitalism is boiling it slowly (poor
| people in first world countries might not afford a dentist
| but they're not starving yet).
|
| We need a system that rewards work - human time and
| competence.
|
| There are really only 2 resources in the world - natural
| resources and human time. Everything else is built on top
| of those. And the people providing their time should be
| rewarded, not those who are in positions of power which
| allow them to extract value while not providing anything in
| return.
| martin-t wrote:
| 56 minutes, 4 downvotes, HN is truly full of temporarily
| embarrassed millionaires.
|
| Does anybody here really think rich people deserve to
| just get richer faster than any working person can? Does
| anybody really believe that buying up homes and companies
| and raking in money for doing absolutely nothing is what
| we should be rewarding?
|
| Then put your name behind it.
| Legend2440 wrote:
| The Miyzaki quote is out of context, he isn't talking about
| generative AI but rather a 2016 animation of a creepy
| zombie whose limbs are controlled by AI.
| martin-t wrote:
| While this is true, it's hard to imagine people spending
| years perfecting the style would be happy to see it
| copied effortlessly without any compensation while people
| who made the copying possible are rolling in cash.
|
| This is not just about copyright infringement or
| plagiarism.
|
| Automatically generating text, images and videos based on
| training data and a tiny prompt is fundamentally about
| taking someone's work and making money off of it without
| giving anything in return.
| mattgreenrocks wrote:
| Don't worry, I'm sure someone will roll up and claim that
| it's just "democratization" of that style and the prompt
| authors exhibit as much creativity as the artists
| themselves.
| squidsoup wrote:
| No, the zombie context is actually not that relevant,
| given he says "We as humans are losing faith in
| themselves" in response to the AI animation. He's clearly
| disgusted by the entire concept of machine generated art.
| larodi wrote:
| Indeed is difficult to NOT share this resentment, should
| anyone understand what actually happens.
| martin-t wrote:
| People are willingly blind.
|
| Kids are happy that homework takes less time. Teachers
| are happy that grading the generated homework takes less
| time. Programmers are happy they can write the same
| amount of code in less time. Graphic designers are happy
| they can get an SVG from a vague description immediately.
| Writers are happy they can generate filler from a few
| bullet points quickly.
|
| But then someone comes along, notices people are not
| working most of the time, fires three quarters of them
| and demands 4x increased output from the rest. And they
| can do it because the "AI" is helping them.
|
| Except they don't get paid any more. The company makes
| the same amount of money for less cost.
|
| So where does the difference go? To the already rich who
| own the company and the product.
| minimaxir wrote:
| This is interesting because every recent model demos
| _conspiciously_ avoids using IP in their demo examples for
| obvious reasons.
| aubanel wrote:
| That, and the dragon looking straight out of How to Train Your
| Dragon - I wonder if they have agreements with the right
| holders, or if they expect massive lawsuits to create free
| advertising for their launch.
| beernet wrote:
| Overall, appears rather underwhelming. Long way to go still for
| video generation. Also, launching this as a social app seems like
| yet another desperate try to productize and monetize their tech,
| but this is the position big VC money forces you into.
| falcor84 wrote:
| I've perhaps been away from the scene for a bit, but I'm very
| impressed. To me this is absolutely "video generation", and I
| don't get your disdain for productization and monetization;
| last I checked this wasn't "Basic Research News".
| DetroitThrow wrote:
| I don't think it's disdainful to point out the lack of PMF
| for a dedicated app for Sora, nor how its behind competitors
| who don't require a dedicated social app. No need to strawman
| the guy, I think it's okay to be reasonably critical of ideas
| still on this website.
|
| Inb4 make your own video model and see how easy it is
| msp26 wrote:
| The voice quality in the generated vids is surprisingly awful.
| gmueckl wrote:
| That's the first thing I noticed, too. The first words you hear
| in the trailer sounds like someone ran the voice through a comb
| filter. It's so bad it made my skin crawl immediately.
| minimaxir wrote:
| OpenAI apparently assumes that the primary users of Sora 2/the
| Sora app will be Gen Z, especially with the demo examples shown
| in the livestream. If they are trying to pull users from TikTok
| with this, it won't work: there's _some_ nuance to Gen Z
| interests than being quirky and random, and if they did indeed
| pull users from TikTok then ByteDance could easily include their
| own image /video generators.
|
| Sora 2 itself as a video model doesn't seem better than Veo
| 3/Kling 2.5/Wan 2.2, and the primary touted feature of having a
| consistent character can be sufficiently emulated in those models
| with an input image.
| usaar333 wrote:
| Physics seems better than veo 3 at least from demo videos
| bflesch wrote:
| Good point. I think OpenAI lacks the cultural understanding
| that tiktok is providing their users not only with
| entertainment but also social things like trends, reviews,
| gossip, self-expression. These aspects are not included in the
| sora experience.
| rhetocj23 wrote:
| This is going to sound crass but idc - OAI is just full of
| geeks, when what is needed is people who are more akin to
| hippies - thats pretty much what Apple was in the early days.
|
| Its no use building technology when its not married with the
| humanities and liberal arts.
| mdrzn wrote:
| If this is anything near the demo they have been released, this
| seems incredibly good at physics. Wow. Can't wait to try the new
| app.
| jsheard wrote:
| Sora 1 was also lauded as being incredibly good at physics
| based on the early cherry-picked examples. The phrase "world
| simulator" was thrown around a lot. That didn't last long once
| people finally got their hands on it though.
| DetroitThrow wrote:
| The space dog and ice skater demo make it seem still very
| close to Sora 1
| benjiro wrote:
| Kind of wondersome if they will start to combine LLM
| generation with actual world models/GPU engines. Imagine that
| your model generates the wireframes, the Engine generates the
| physics and then another model fills in the actual visuals,
| and gaps... So you have realistic physics and gaps are filled
| in... Will also help with image retention more, if objects
| moved behind each other.
| techpression wrote:
| The demo on their homepage shows really bad physics. There's a
| lot of it, but that doesn't mean it's correct. The hair of Sam
| looks like a paper cutout in almost every shot.
| spaceman_2020 wrote:
| Kling 2.5 is already pretty good at physics
|
| I don't expect Sora2 to be SOTA. The Chinese models are further
| ahead in video/image gen
| fariszr wrote:
| Did they make human voices sound robotic on purpose? Is that some
| kind of Ai fingerprinting? It's way too obvious
| minimaxir wrote:
| It's very hard for simultaneous good audio generation with
| video generation (simultaneous generation is necessary to
| maintain lip sync). Veo 3 et al also have flat monochannel
| audio, but not as bad as these Sora 2 demos.
| causal wrote:
| IDK if the site is being hugged to death but I can only load the
| first video. Even in just one viewing there were noticeable
| artifacts, so my impression is that Veo is still in the lead
| here.
| qafy wrote:
| Yeah I am curious what the actual resolution of these videos
| will be. The launch videos on this link will only play in like
| 360p for me.
| S0und wrote:
| I find it comical that OpenAI with all the power of CharGPT even
| them are unable to release an app for both iOS and Android at the
| same time. Wow, good marketing for Codex.
| aizk wrote:
| That is more of a statement of the complete dominance of
| iPhones among gen z.
| bigyabai wrote:
| Or Sama's documented reverence for Apple products. We _are_
| talking about the guy who sold Tim Cook his AI for $0.00, he
| 's not exactly got the horse drawing the cart here.
| drexlspivey wrote:
| Google sold Tim Cook their search engine for $-25B per year
| gmuslera wrote:
| Not even for all regions for iOS
| rd wrote:
| https://apps.apple.com/us/app/sora-by-openai/id6744034028
|
| App link
|
| edit: CBN80W for an invite code
| throwup238 wrote:
| I downloaded the app but I get a "Sora is invite only" screen
| after logging in to my OpenAI account and asking for an invite
| code.
| Tiberium wrote:
| > You can sign up in-app for a push notification when access
| opens for your account.
|
| You need to be in the US/Canada and wait for this
| notification, and when you get an invite you can start using
| it in the app and on sora.com. And apparently you get 4 more
| invite codes that you can share with anyone, e.g. Android
| users:
|
| > Android users will be able to access Sora 2 via
| http://sora.com once you have an invite code from someone who
| already has access
| qingcharles wrote:
| It's wild that I have a paid account but I have to scour
| the Internet to find someone else with a paid account and
| beg them for an invite code to use the product I already
| paid for. Make it make sense.
| gretch wrote:
| One thing that would make sense is for you to not pay any
| more.
|
| But if you do, that signals to the company this is all
| perfectly okay.
| Y_Y wrote:
| Do you really want a "social" app for a firehose of high-
| fidelity slop?
| solfox wrote:
| This access code is "no longer available" :(
| cactusplant7374 wrote:
| Check the browser console. The endpoint is returning 429 for
| me. So it might not even be accepting codes depending on how
| many you try.
| DetroitThrow wrote:
| Just seeing the examples that I assumed are cherry picked, it
| seems like they're still behind on Google when it comes to video
| generation, the physics and stylized versions of these shots seem
| not great. Veo3 was such a huge leap and is still ahead of many
| of the other large AI labs.
| rushingcreek wrote:
| The most interesting thing by far is the ability to include video
| clips of people and products as a part of the prompt and then
| create a realistic video with that metadata. On the technical
| side, I'm guessing they've just trained the model to
| conditionally generate videos based on predetermined characters
| -- it's likely more of a data innovation than anything
| architectural. However, as a user, the feature is very cool and
| will likely make Sora 2 very useful commercially.
|
| However, I still don't see how OpenAI beats Google in video
| generation. As this was likely a data innovation, Google can
| replicate and improve this with their ownership of YouTube. I'd
| be surprised if they didn't already have something like this
| internally.
| visarga wrote:
| > the ability to include video clips of people and products as
| a part of the prompt and then create a realistic video with
| that
|
| This is something I would not like to see, I prefer product
| videos to be real, I am taking a risk with my money. If the
| product has hallucinated or unrealistic depiction it would be a
| kind of fraud.
| BeetleB wrote:
| I believe existing laws already cover that issue.
| pton_xd wrote:
| Someone remind me the benefits of mass produced fake videos
| again?
| ToucanLoucan wrote:
| - Political propaganda
|
| - Scamming people at scale
|
| - Nonconsensual pornography
|
| - Juicing engagement metrics for fading social media sites
|
| - The ongoing destruction of truth as a concept in our
| increasingly atomized and divided world
| jablongo wrote:
| I think the last one takes the cake.
| chis wrote:
| I imagine it's incredibly useful for prototyping movies, tv,
| commercials before going to the final version. CGI will
| probably get way cheaper too with some hybrid approach.
|
| Obviously this will get used for a lot of evil or bad as well
| greyk47 wrote:
| can you imagine a billion dollar company promoting their new
| pre-vis app?
| jsheard wrote:
| I feel like that's missing the point of pre-vis anyway, its
| purpose is to lay down key details with precision but
| without regard for fidelity (e.g.
| https://youtu.be/KMMeHPGV5VE), a system with high fidelity
| but very loose control is the exact opposite of what they
| want.
| minimaxir wrote:
| It's fun: maybe not for everyone, but there's clearly
| sufficient interest in it.
|
| Whether said fun is "worth" the social and economic costs is a
| separate issue.
| IncreasePosts wrote:
| I can have an idea and see a video of something like my idea
| pretty quickly.
|
| What are the benefits of what you do? Does anyone know?
| jamiecurle wrote:
| Fun.
| observationist wrote:
| ... how dare you, sir. That is entirely unacceptable and you
| will be reported to the ministry of proper living!
|
| Regardless of the slop, some people will learn to use it
| well. You have stuff like NeuralViz - quite the sight! - and
| other creators will follow suit, and figure out how to use
| the new tools to produce content that's worth engaging with.
| Bigfoot vlogs and dinosaur chase scenes, all that stuff is
| mostly just fun.
|
| People like to play. Let them play. This stuff looks fun, and
| beats Sora 1 by a long shot.
|
| Hopefully it catalyzes
| jasonsb wrote:
| Democracy? Strengthened! Nothing says "informed electorate"
| like not knowing if a politician actually said they support
| nazism or if it was just a hyper-realistic AI puppet.
|
| Trust in media? Soaring! Why believe your eyes or ears when you
| can doubt everything equally?
|
| Journalism? Thriving! Reporters now get to spend their days
| playing forensic video detective instead of, you know,
| reporting news.
|
| Social harmony? Better than ever! Nothing brings people
| together like shared paranoia and the collective shrug of "I
| guess truth is dead now."
|
| Honestly, what could possibly go wrong?
| theLiminator wrote:
| lol i wonder if this will create a market for PKI at the
| image sensor level so that videos will be cryptographically
| signed and baked into the actual video stream with
| steganography.
| bsenftner wrote:
| Advertising: you (her) wearing new clothing before purchase,
| hair/glasses/makeup, make overs; guys after 3 months of gym
| membership, you driving the new car, you in this specific new
| home... etc, etc... I'm surprised this is not already
| everywhere, but people are too occupied making nsfw and fantasy
| violence clips.
| a2128 wrote:
| Targeted advertising has become just manipulation. I don't
| know if personalized advertisement videos for everyone
| promoting a fake world that doesn't exist is really a benefit
| for the world...
| bsenftner wrote:
| If course it's not a benefit, but it's an advertising angle
| that will work very well with a class of gullible
| consumers, and that is enough to justify it being plastered
| everywhere. I don't write these rules, I just notice them.
| thorum wrote:
| People are doing cool things with it. Here's one example:
|
| https://www.tiktok.com/@dreamrelicc
|
| Before AI, each video on this channel would have taken a large
| team with a Hollywood budget to create. In a few more years,
| one person may be able to turn their creative vision into a
| full-length movie.
| j4hdufd8 wrote:
| What are the benefits of those videos?
| minimaxir wrote:
| What are the benefits of producing any video?
| drexlspivey wrote:
| What are the benefits of this comment?
| j4hdufd8 wrote:
| Challenging the value of AI generated "art"
| frde_me wrote:
| Then the purpose of those videos is to challenge the
| value of non AI generated "art"
|
| (half sarcastic, but you could make the argument that
| most art has no benefit besides to the person that made
| the art)
| j4hdufd8 wrote:
| Nice! I enjoyed this sub thread. I'm not sure what I
| conclude but I enjoyed thinking about this.
| busymom0 wrote:
| > People are doing cool things with it
|
| Things are cool because they are unique, very hard to create,
| and require creativity. When those things become cheap
| commodities, they are no longer cool.
| minimaxir wrote:
| The same could be said about software, and it's safe to say
| that open-source software making complex workflows easier
| and more efficient is a net good.
|
| Making better tools is better for everyone: the median
| usage of those tools downstream is a separate issue.
| viccis wrote:
| If you're comparing how art is evaluated to how software
| is evaluated then it sounds like you only understand one
| or the other.
| cubefox wrote:
| Indeed. Art is partially evaluated by how impressive it
| is. That's why posting AI images on social media won't
| yield a lot of likes anymore. People have gotten used to
| images being easy to create, so they aren't seen as
| valuable anymore. The same will be true for videos.
|
| AI pictures today are much less impressive than Dall-E 2
| pictures were a few years ago, despite the fact that the
| models are much better nowadays. Currently AI videos can
| still be impressive, but this will quickly become a thing
| of the past.
|
| Then people will move from trying to create art to
| creating "content". That is, non-artistic slop.
| Advertisements. Porn. Meme jokes. Click bait. Rage bait.
| Propaganda. Etc.
| thorum wrote:
| I would argue that we just get pickier and more sensitive
| to slop. When everyone can make a movie, the standard for a
| good movie will be higher. Many current Hollywood films
| wouldn't make the cut. But maybe some kid in Nigeria makes
| the greatest film of all time.
| cubefox wrote:
| By that logic, some kid in Nigeria could have written the
| greatest book of all time. At least by commonly accepted
| measures, that didn't happen.
| squidsoup wrote:
| Hard to interpret that comment as anything but racist.
| Chinua Achebe is widely considered one of the greatest
| modern novelists. He was 28 when he wrote Things Fall
| Apart.
| rhetocj23 wrote:
| This is absolutely horrible.
|
| People need to be exposed to what is real. Not more
| artificial stuff.
|
| I think this is the point at which humanity will finally puke
| and reject this crap.
|
| Just because a small segment of people like it doesnt mean
| the mass majority will.
| qingcharles wrote:
| Maybe your real is good. For most people on Earth, real
| isn't that great.
| FergusArgyll wrote:
| I personally love Monet, he's not for everyone, I know, but
| I'm sure you can find some art you appreciate
| cubefox wrote:
| You probably don't personally love AI generated
| impressionist content.
| FergusArgyll wrote:
| No, but there's some stuff that are really creative.
| Ironically I think the reason I'm more positive about it
| is because I only encounter AI generated (non-text media)
| ~ once a week / 2 weeks.
| jasonsb wrote:
| > In a few more years, one person may be able to turn their
| creative vision into a full-length movie.
|
| Yes, but at the same time the value of video production will
| quickly drop to 0. Or to whatever it costs to generate that
| video in terms of tokens.
| intended wrote:
| The value will shift to search or curation - if the cost to
| produce drops to nil, then the value will be in finding good
| content amongst a flood of sameness.
| asdev wrote:
| I speak for everyone when I say we don't need these videos at
| all and would be better off without them
| thorum wrote:
| I disagree, so not everyone, I guess!
| knowaveragejoe wrote:
| I love the aesthetic in this person's videos, I just wish it
| wasn't on tiktok :(
| squidsoup wrote:
| The problem is, it isn't their aesthetic, it's a
| resynthesis of the aesthetic of someone else's work.
| 2OEH8eoCRo0 wrote:
| Can it generate an analog clock displaying a given time?
| martypitt wrote:
| Even if it can't, that wouldn't make this demo any less
| impressive.
| gvv wrote:
| Any idea if or when it will be available in EU?
| https://apps.apple.com/us/app/sora-by-openai/id6744034028
|
| edit: as per usual it's not yet...
| aaroninsf wrote:
| Someone who doesn't follow the moving edge would be forgiven for
| being confused by the dismissive criticism dominating this thread
| so far.
|
| It's not that I disagree with the criticism; it's rather that
| when you live on the moving edge it's easy to lose track of the
| fact that things like this are _miraculous_ and I know not a
| single person who thought we would get results "even" like this,
| this quickly.
|
| This is a forum frequented by people making a living on the edge
| --get it. But still, remember to enjoy a little that you are
| living in a time of miracles. I hope we have leave to enjoy that.
| cubefox wrote:
| Yeah. Just a few years ago, people here would have said stuff
| like that was decades away at best and pure science fiction at
| worst.
| qoez wrote:
| I know the comments here are gonna be negative but I just find
| this so sick and awesome. Feels like it's finally close to the
| potential we knew was possible a few years ago. Feels like a
| pixar moment when CG tech showed a new realm of what was possible
| with toy story
| m3kw9 wrote:
| No doubt they can create Hollywood quality clips if the tools
| are good enough to keep objects consistent, example, coming
| back to the same scene with same decor and also emotional
| consistency in actors
| gretch wrote:
| > keep objects consistent
|
| I think this is not nearly as important as most people think
| it is.
|
| In hollywood movies, everyone already knows about "continuity
| errors" - like when the water level of a glass goes up over
| time due to shots being spliced together. Sometimes shots
| with continuity errors are explicitly chosen by the editor
| because it had the most emotional resonance for the scene.
|
| These types of things rarely affect our human subjective
| enjoyment of a video.
|
| In terms of physics errors - current human CGI has physics
| errors. People just accept it and move on.
|
| We know that superman can't lift an airplane because all of
| that weight on a single point of the fuselage doesn't hold,
| but like whatever.
| inerte wrote:
| It all depends on quantity and "quality" of the continuity
| errors. There's even a job for it
| https://en.wikipedia.org/wiki/Script_supervisor
| ileonichwiesz wrote:
| Water level in a glass changing between shots is one thing,
| the protagonist's face and clothes changing is another.
| bbor wrote:
| Well put. Honestly the actor part is mostly solved by
| now, the tricky part is depicting any kind of believable,
| persistent space across different shots. Based off of
| amateur outputs from places like
| https://www.reddit.com/r/aivideo/, at least!
|
| This release is clearly capable of generating mind-
| blowingly realistic short clips, but I don't see any
| evidence that longer, multi-shot videos can be automated
| yet. With a professional's time and existing editing
| techniques, however...
| echelon wrote:
| Location consistency is important. Even something as
| simple and subtle as breaking the 180-rule [1] feels
| super uncanny to most audiences. Let alone changing the
| set the actor occupies, their wardrobe, props, etc.
|
| There are lots of tools being built to address this, but
| they're still immature.
|
| https://x.com/get_artcraft/status/1972723816087392450
| (This is something we built and are open sourcing - still
| has a ways to go.)
|
| ComfyUI has a lot of tools for this, they're just hard to
| use for most people.
|
| [1] https://en.wikipedia.org/wiki/180-degree_rule
| cryptoz wrote:
| I wonder if this stuff is trained on enough Hallmark movies
| that even AI actors will buy a hot coffee at a cafe and
| then proceed to flail the empty cup around like the humans
| do. Really takes me out of the scene every time - they
| can't even put water in the cup!?
| layer8 wrote:
| People got used to James Bond actors changing between
| movies, but from scene to scene in the same movie would be
| a bit confusing.
| beefnugs wrote:
| No way man, this is why i loved Mr Robot, they actually
| payed a real expert and worked story around realism and not
| just made up gobbleygook that shuts my brain off entirely
| to its nonsense
| loudmax wrote:
| These videos are a very impressive engineering feat. There are
| a lot of uses for this capability that will be beneficial to
| society, and in the coming years people will come up with more
| good uses nobody today has thought of yet.
|
| But clearly we also see some major downsides. We already have
| an epidemic of social media rotting people's minds, and
| everything about this capability is set to supercharge these
| trends. OpenAI addresses some of these concerns, but there's
| absolutely no reason to think that OpenAI will do anything
| other than what they perceive as whatever makes them the most
| money.
|
| An analogy would be a company coming up with a way to
| synthesize and distribute infinite high-fructose corn syrup.
| There are positive aspects to cheaply making sweet tasting
| food, but we can also expect some very adverse effects on
| nutritional health. Sora looks like the equivalent for the
| mind.
|
| There's an optimistic take on this fantastic new technology
| making the world a better place for all of us in the long run,
| after society and culture have adapted to it. It's going to be
| a bumpy ride before we get there.
| kimbler wrote:
| I actually wonder if this will kill off the social apps and
| the bragging that happens. It will be flooded by people
| faking themselves doing the unimaginable.
| artursapek wrote:
| This is also my thesis. The internet is going to be
| saturated with AI slop indiscernible from real content.
| Once it reaches a tipping point, there will no longer be
| much of a reason to consume the content at all. I think
| social networks that can authenticate video/photo/text
| content as human-created will be a major trend in a few
| years.
| Mariehane wrote:
| But then you're creating an incentive for the AI slop to
| become so realistic it is indistinguishable from actual
| video.
|
| Unless there some fundamental, technical way to
| distinguish the two, I wonder who would win?
| artursapek wrote:
| there would need to be cameras that can cryptographically
| sign videos with trusted vendor keys, or perhaps there is
| some other solution.
| fabrice_d wrote:
| This is what https://c2pa.org/ is for. I think some
| camera vendors already have support.
| sigbottle wrote:
| I regularly get AI movie recaps on my shorts and I just
| eat it up.
|
| The very fact that I (or billions of others) waste time
| on shorts is an issue. I don't even play _games_ anymore,
| it 's just shorts. That is a concerning rewiring of the
| brain :/
|
| Guess what I`m trying to say is that, there is a market
| out there. It's not pretty, but there certainly is.
|
| Will keep trying to not watch these damn shorts...
| sumeruchat wrote:
| there will be billions of people consuming the content
| larodi wrote:
| Depending on which internet you do mean, cause meta &
| insta are NOT THE Internet.
| layman51 wrote:
| I have no clue if the reactions are real, but there are
| some videos online of people showing their grandparents
| gameplay from Grand Theft Auto games trying to convince
| them that it is real footage. The point of the videos is
| to laugh at their reactions where they question if it
| really happened, etc.
|
| Maybe this will result in something similar, but it can
| affect more people who aren't as wary.
| hsuduebc2 wrote:
| Heh, fast forward a few years and nobody's surprised
| anymore when someone falls for a video which is the
| result of two sentences long instruction.
| dvngnt_ wrote:
| Right now with kids, the current trend is to prank their
| parents using Gemini into thinking they let a homeless
| guy in their house
|
| https://www.tiktok.com/discover/ai-homeless-people-in-my-
| hou...
| marcosdumay wrote:
| Yes, I wonder if the content distribution networks that
| call themselves "social networks" can even survive
| something like this.
|
| Of course, the ones focusing on the content can always
| editorialize the spam out. And in real social networks you
| ask your friends to stop making that much slop. But this
| can be finally the end of Facebook-like stuff.
| shoobiedoo wrote:
| > There are a lot of uses for this capability that will be
| beneficial to society
|
| Please enlighten me. What are they? If my elderly grandma is
| on her deathbed and I have no way to get to see her before
| she passes, will she get more warmth and fond memories of me
| with a clip of my figure riding an AI generated dragon saying
| goodbye, or a handwritten letter?
| latexr wrote:
| > There are a lot of uses for this capability that will be
| beneficial to society
|
| Are there? "A lot" of them? Please name a few that will be
| more beneficial than the very obvious detrimental uses like
| "making up life-destroying lies about your political
| opponents or groups of people you want to vilify" or "getting
| away with wrongdoing by convincing the judge a real video of
| yourself is a deepfake".
|
| That last one has already ben tried, by the way.
|
| https://www.theguardian.com/technology/2023/apr/27/elon-
| musk...
| croes wrote:
| I already get enough AI spam and scam videos on social media. I
| don't need them to be better quality
| Scrapemist wrote:
| Pixar moment for me means a novel techonology evoking a
| profound emotional response for the first time. This was not
| it.
| linuxftw wrote:
| The ability for the masses to create any video just by
| typing, among the other features, is not novel technology? Or
| is it just the lack of emotional response?
| varispeed wrote:
| I still feel this is limited by what it learned from. It looks
| cool but it also looks like something I'd dreamt or saw
| flicking through TV channels. Kind of like spam for the eyes.
| colesantiago wrote:
| > Feels like a pixar moment when CG tech showed a new realm of
| what was possible with toy story
|
| @qoez
|
| > The first entirely AI generated film (with Sora or other AI
| video tools) to win an Oscar will be less than 5 years away.
|
| https://news.ycombinator.com/item?id=42368951
|
| This prediction of mine was only 10 months ago.
|
| Imagine when we and if we get to 5 years.
| hansmayer wrote:
| Potential for what exactly? More of 30-sec slop?
| jsnell wrote:
| Doing this as a social app somehow feels really gross, and I
| can't quite put to words why.
|
| Like, it should be _preferable_ to keep all the slop in the same
| trough. But it 's like they can't come up with even one
| legitimate use case, and so the best product they can build
| around the technology is to try to create an addictive loop of
| consuming nothing but auto-generated "empty-calories" content.
| pavon wrote:
| I see it more as recognizing that it will take time to be good
| enough for other use cases so for this release they are
| targeting it as just something to have fun with. After seeing
| LLMs crammed into everything whether it makes sense or not, I
| can appreciate that.
| dweekly wrote:
| So a social network that's 100% your friends doing silly AI
| things?
|
| I feel like this is the ultimate extension of "it feels like my
| feed is just the artificial version of what's happening my
| friends and doesn't really tell me anything about how they're
| actually faring."
| al_borland wrote:
| Social media also tends to highlight the best parts of people's
| lives, creating unrealistic expectations and views for those
| consuming it and looking at their real life. Now social media
| won't even be a highlight reel, but completely fabricated.
|
| I have to imagine there will be a rebellion against all of this
| at some point, when people simply can't take the false
| realities anymore. What is the alternative? Ready Player One?
| The Matrix? Wall-E?
| m3kw9 wrote:
| Which seem to level the play field, at least virtually
| al_borland wrote:
| Maybe inside of a social network specially for AI, but a
| concerning number of people don't realize images and videos
| are AI, even when it's bad AI. As it gets better, and
| starts integrating the poster's image (like Sora 2), that's
| going to get even worse.
| kjs3 wrote:
| Some people use filters/photoshop to artificially juice their
| images; now they can use AI to artificially juice every aspect
| of their on-line presence.
| doctorhandshake wrote:
| I built an MVP of this [1] with images (not video) and in more
| of an Instagram style (not tiktok) back in '22, with the
| tagline 'What if social media were literally fake?'
|
| I am bullish on this, albeit with major concerns in many
| domains. It was fun and addictive as hell with images. With
| video it will be wild.
|
| [1] https://hardwork.party/cheese/
| superfrank wrote:
| I just watched the announcement video and something about it
| just gives me the ick. The whole time I just had the uncanny
| valley feeling.
|
| The technology itself is super impressive, but a social media
| app of AI slop doesn't feel like the best use of it. I'm old
| enough to not really be interested in social media in general
| anymore, so maybe I'm just out of touch, but I just can't see
| this catching on. It feels like the type of thing that people
| will download, use a few times until the novelty wears off and
| then never open again.
| jpalomaki wrote:
| Sounds like a way to make everybody aware of Sora and its
| capabilities.
|
| I bet the real goal is to make money from long tail of
| corporate market ( ads, info videos etc).
| taytus wrote:
| Honest question: What problem does this solve?
| lawlessone wrote:
| Liberation of employers from the shackles of their employees
| martypitt wrote:
| Thing that was previously very expensive, manual and took a
| long time to do, and is done A LOT, is now made faster and
| cheaper by computers.
|
| Pretty much the same problem we all work on every day in
| $DAY_JOB.
| nextworddev wrote:
| Fun and games until someone uses a tool like this to scam
| your family
| lawlessone wrote:
| It's ok, they're making the market for anti-ai tools much
| much bigger. (whether those tools work or not is a
| different issue)
| smith7018 wrote:
| OpenAI needing something to show to investors to say "See, this
| is why we need $1T."
| andybak wrote:
| What problem does what solve? Video generation models in
| general or Sora 2 specifically?
| squidsoup wrote:
| It facilitates the generation of political propaganda.
| mempko wrote:
| It's obvious there is no way OpenAI can keep videos generated by
| this within their ecosystem. Everything will be fake, nothing
| real. We are going to have to change the way we interact with
| video. While it's obviously possible to fake videos today, it
| takes work by the creator and takes skill. Now it will take no
| skill so the obvious consequence of this is we can't believe
| anything we see.
|
| The worst part is we are already seeing bad actors saying 'I
| didn't say that' or 'I didn't do that, it was a deep fake'. Now
| you will be able to say anything in real life and use AI for
| plausible deniability.
| mmmrtl wrote:
| I think that's the point... Then world coin comes to the rescue
| roxolotl wrote:
| World coin is so delightfully dystopian. You could drop it
| wholesale into a superhero movie and it would be believable
| as the supervillain's plot.
| kjs3 wrote:
| _We are going to have to change the way we interact with
| video._
|
| I doubt it will be for the better. The ubiquity of AI deepfakes
| just reenforces entrenchment around "If the message reinforces
| my preconceived notion, I believe it and think anyone who calls
| it fake is stupid/my enemy/pushing an agenda. If the message
| contradicts my preconceived notion, it's obviously fake and
| anyone who believes it is stupid/my enemy/pushing an agenda.".
| People don't even take the time to think "is this even
| _plausible_ ", much less do the intellectual work to verify.
| armchairhacker wrote:
| Record things with 2 cameras.
|
| Today's Sora can produce something that resembles reality from
| a distance, but if you look closely, especially if there's
| another perspective or the scene is atypical, the flaws are
| obvious.
|
| Perhaps tomorrow's Sora will overcome the the "final 10%" and
| maintain undetectable consistency of objects in 2 perspectives.
| But that would require a spatial awareness and consistency that
| models still have a lot of trouble with.
| gdulli wrote:
| It's also possible we remain stuck in the uncanny valley
| forever, or at least for the rest of our lives.
|
| It's possible to produce _some_ video or image that looks real,
| cherry-picked for a demo, but not possible to produce any
| arbitrary one you want that will end up passable.
| SV_BubbleTime wrote:
| >Everything will be fake, nothing real. We are going to have to
| change the way we interact with video.
|
| I'm optimistic here.
|
| Look at 1900s tech like social security number/card, and paper
| birth certificates. Our world is changing and new systems of
| verification will be needed.
|
| I see this as either terribly dystopian - or - a possibility
| for the mass expansion of cryptography and encrypted/signed
| communication. Ideally in privacy preserving ways because
| nothing else will make as much sense when it comes to the
| verification that countries will need to give each other even
| if they want backdoor registry BS for the common man.
|
| Breaking changes get fixes.
| mike_hearn wrote:
| It's not that obvious. iOS is pretty secure, if they keep the
| social network and cameo feature limited to that there might
| not be good ways to export videos off the platform onto others
| beyond pointing a camera at the tablet screen. And beyond there
| being lots of ways to watermark stuff to be detectable, nothing
| stops the device using its own camera to try and spot if it's
| being recorded. The bar can be raised quite high as long as
| you're willing to exclude any device that isn't an iPhone/iPad.
| whimsicalism wrote:
| Find this sort of innovation far less interesting or exciting
| than the text & speech work, but it seems to be a primary driver
| of adoption for the median person in a way that text capability
| simply is not.
| liuliu wrote:
| Video generation is extremely exciting a.k.a. https://video-
| zero-shot.github.io/
|
| However, personalization (teleporting yourself into a video
| scene) is boring to me. At its core, it doesn't generate new
| experience to me. My experience is not defined by photos /
| videos I took on a trip.
| currymj wrote:
| I also can't think of a reason why I would ever want to look at
| an AI generated video.
|
| however as they hint at a little in the announcement, if video
| generation becomes good enough at simulating physics and
| environments realistically, that's very interesting for
| robotics.
| jablongo wrote:
| Sam Altman has made (for me) encouraging statements in the past
| about short-form video like TikTok being the best current example
| of misaligned AI. While this release references policies to
| combat "Doomscrolling and RL-sloptimization", it's curious that
| OpenAI would devote resources to building a social app based on
| AI generated short form video, which seems to be a core problem
| in our world. IMO you can't tweak the TikTok/YouTube shorts
| format and make it a societal good all of a sudden, especially
| with exclusively AI content. This is a disturbing development for
| Altman's leadership, and sort of explains what happened in 2023
| when they tried to remove him... -> says one thing, does the
| opposite.
| bigyabai wrote:
| Sam Altman is a businessman. His job is to say whatever
| assuages his market, and that includes gaslighting you when
| you're disgusted by AI.
|
| If you never expected Altman to be the figurehead of principled
| philosophy, none of this should surprise you. _Of course_ the
| startup alumni guy is going to project maligned expectations in
| the hopes of being a multi-trillion dollar company. The
| shareholders love that shit, Altman is applying the same
| lessons he learned at Worldcoin to a more successful business.
|
| There was never any question _why_ Altman was removed, in my
| mind. OpenAI outgrew it 's need for grifters, but the grifter
| hadn't yet outgrown his need for OpenAI.
| estearum wrote:
| > His job is to say whatever assuages his market
|
| I understand the cynicism but this is in fact _not_ the job
| of a businessman. We shouldn 't perpetuate the pathological
| meme that it is.
| bnop wrote:
| So the job of a businessman is not to increase shareholder
| value?
| jablongo wrote:
| To be clear I'm not disgusted by AI in general, I'm disgusted
| by short form video and AI/ML in service of dopamine reward
| loop hacking.
| pants2 wrote:
| I'm optimistic about the Sora app! My hope is that it becomes
| much more whimsical and fun than TikTok because everyone on the
| app knows that all content is fake. Hopefully that means less
| rage-bait and more creative content, like OG YouTube. Nobody's
| going to get their news from Sora because it's literally 100%
| fake.
| lxgr wrote:
| > it becomes much more whimsical and fun than TikTok because
| everyone on the app knows that all content is fake.
|
| Sounds about as plausible as "ironically taking heroin".
|
| > Nobody's going to get their news from Sora because it's
| literally 100% fake.
|
| I'm with Neal Stephenson ("Fall", in this case) on this
| prediction, although I really hope I'm wrong.
| lxgr wrote:
| That said... does anyone have an invite code?
| jablongo wrote:
| Why would it be more like OG YouTube, when the content they
| demoed very closely resembles YouTube shorts? The key
| difference is OG YouTube was long form.
| bonoboTP wrote:
| > much more whimsical and fun than TikTok
|
| In the early years everyone told me that TikTok is actually
| fun and whimsical (like just after it stopped being
| musical.ly), and it's all about fun collaboration, and
| amateur comedy sketches, fun dances and lipsyncs, and people
| posting fun reactions to each other etc, all lighthearted and
| that social media is finally fun again!
| xeeeeeeeeeeenu wrote:
| >IMO you can't tweak the TikTok/YouTube shorts format and make
| it a societal good all of a sudden, especially with exclusively
| AI content.
|
| I agree. At best, short videos can be entertainment that
| destroys your attention span. Anything more is impossible. Even
| if there were no bad actors producing the content, you can't
| condense valuable information into this format.
| modeless wrote:
| I can see it being interesting to create wacky fake videos of
| your friends for a week or two, but why would people still be
| using this next year?
|
| I watch videos for two reasons. To see real things, or to consume
| interesting stories. These videos are not real, and the
| storytelling is still very limited.
| derac wrote:
| I'm no Nostradamus, but I predict these models will be much
| better in a year.
| pr337h4m wrote:
| soft porn
| qingcharles wrote:
| You only watch real things? Have you never watched a movie?
| modeless wrote:
| > or to consume interesting stories
| FergusArgyll wrote:
| In the right hands it's a new art medium. Some (few, maybe)
| midjourney generations are serious art.
|
| So, for the same reason you'd go to a local art gallery
| mempko wrote:
| It's obvious there is no way OpenAI can keep videos generated by
| this within their ecosystem. Everything will be fake, nothing
| real. We are going to have to change the way we interact with
| video. While it's obviously possible to fake videos today, it
| takes work by the creator and takes skill. Now it will take no
| skill so the obvious consequence of this is we can't believe
| anything we see.
|
| The worst part is we are already seeing bad actors saying 'I
| didn't say that' or 'I didn't do that, it was a deep fake'. Now
| you will be able to say anything in real life and use AI for
| plausible deniability.
|
| I predict a re-resurgence in life performances. Live music and
| live theater. People are going to get tired of video content when
| everything is fake.
| minimaxir wrote:
| The Sora 2 livestream indicates that videos exported from the
| app will have visual watermarks.
| ileonichwiesz wrote:
| Sure, then you just pump it through another model that
| removes watermarks.
| saltyoldman wrote:
| Let the sloppification of all children's minds begin!!!!!!
| jablongo wrote:
| its well underway already
| password54321 wrote:
| already happening:
| https://x.com/ken_wheeler/status/1954343731579994593
| mempko wrote:
| I predict a re-resurgence in life performances. Live music and
| live theater. People are going to get tired of video content when
| everything is fake.
| nextworddev wrote:
| One would think, but people are spending less on live events
| due to costs
| volkk wrote:
| likely because we haven't yet reached peak slop/exhaustion by
| slop. Soon enough...soon enough
| rvz wrote:
| Buying lots of calls on Live Nation.
| nextworddev wrote:
| Most of human crafted shorts / reels are already slop.
| simonw wrote:
| Anyone with access able to confirm if you can start this with a
| still image and a prompt?
|
| The recent Google Veo 3 paper "Video models are zero-shot
| learners and reasoners" made a fascinating argument for video
| generation models as multi-purpose computer vision tools in the
| same way that LLMs are multi-purpose NLP tools. https://video-
| zero-shot.github.io/
|
| It includes a bunch of interesting prompting examples in the
| appendix, it would be interesting to see how those work against
| Sora 2.
|
| I wrote some notes on that paper here:
| https://simonwillison.net/2025/Sep/27/video-models-are-zero-...
| andrewguenther wrote:
| Yes, you can start with a still and a prompt
| andybak wrote:
| I've got used to immediately checking availability. In this case
| - iPhone app is US + Canada only and the website is invite only.
|
| Going back to sleep. Wake me up when it's available to me.
| outlore wrote:
| in a computer graphics course i took, we looked through how
| popular film stories were tied to the technical achievements of
| that era. for example, toy story was an story born from the new
| found ability to render plastics effectively. similarly, the sora
| video seems to showcase a particular set of slow moving scenes
| (or when fast, disappearing into fluid water and clouds) which
| seem characteristic of this technology at the current moment in
| time
| ChrisArchitect wrote:
| More discussion: https://news.ycombinator.com/item?id=45428122
| dang wrote:
| Comments moved thither. Thanks!
|
| Edit: looks like this post was actually first, so maybe we'll
| reverse the merge
| gorgoiler wrote:
| Impressively high level of continuity. The only errors I could
| really call out are:
|
| 1/ 0m23s: The moon polo players begin with the red coat rider
| putting on a pair of gloves, but they are not wearing gloves in
| the left-vs-right charge-down.
|
| 2/ 1m05s: The dragon flies up the coast with the cliffs on one
| side, but then the close-up has the direction of flight reversed.
| Also, the person speaking seemingly has their back to the
| direction of flight. (And a stripy instead of plain shirt and a
| harness that wasn't visible before.)
|
| 3/ 1m45s: The ducks aren't taking the right hand corner into the
| straightaway. They are heading into the wall.
|
| I do wonder what the workflow will be for fixing any more
| challenging continuity errors.
| fferen wrote:
| Very first frame of the video: green digital text is messed up.
| Stopped watching after that :)
| yoavm wrote:
| The whole pool the ducks are racing at is a completely
| different pool when Sam starts talking.
| cogman10 wrote:
| The snowmobiles were different in each cut. The shape, color,
| and style of the lights were different.
| fwip wrote:
| Not sure if it counts as a continuity error, but in the example
| "Prompt: Martial artist doing a bo-staff kata waist-deep in a
| koi pond", his wooden staff changes shape several times,
| resembling a bow at points. That was the first example I
| noticed as "clearly AI."
| mNovak wrote:
| The Bo staff in the koi pond also seems to involve some
| impossible wrist movements
| tootie wrote:
| The fact that this is their demo to the world and it's full of
| errors implies that average users will only get worse results.
| gorgoiler wrote:
| I'm wary of being that damning, this early. What I want to
| know is, should my video have these kinds of continuity
| errors, how easily can I fix them?
|
| It's ok for this to be a fun toy. (And fun toy while also
| being an astonishing piece of engineering.) But if it wants
| to push beyond fun toy then it would be interesting to see
| how that process works.
|
| Will Sora2 help me sketch out a movie for me, doing 10% of
| the work where I have to reshoot the other 90% for real, or
| will it get me 90% there leaving me only 10% left to do "by
| hand"?
|
| (This is the exact same question, I believe, which is being
| asked of the maintenance burden imposed by vibe coded
| products. They get you 90% then fail spectacularly leaving
| you having to do the bulk of the work again? Or they get you
| 90% of the way and you int have to fill in the gaps to reach
| a stable long term product?)
| cogman10 wrote:
| The video was slam cut together to avoid continuity problems.
| There was a lot of fast camera motion and unconnected scenes.
|
| Particularly bad was the snowmobile sequence. It was literally
| a different snowmobile in every cut.
|
| The racing pool duck scene was a different pool in every shot.
|
| About the only consistent thing was the faces that were spliced
| into the scenes.
|
| I do not really see anything super significant in the demo. It
| looks like this suffers from all the same problems of AI
| generated video. They just hid it by avoiding more then 5
| seconds in the same setting.
| willahmad wrote:
| I wonder about the implications of this tech.
|
| State of the things with doom scrolling was already bad, add to
| it layoffs and replacing people with AI (just admit it, interns
| are struggling competing with Claude Code, Cursor and Codex)
|
| What's coming next? Bunch of people, with lots of free time
| watching non-sense AI generated content?
|
| I am genuinely curious, because I was and still excited about AI,
| until I saw how doom scrolling is getting worse
| m3kw9 wrote:
| I'm wondering how they really prevent uploads of other peoples
| faces if they take a clip of a video of another person. I'm
| sure Apple didn't open up the 3d Face ID scanning to them to
| verify
| pixl97 wrote:
| >What's coming next? Bunch of people, with lots of free time
| watching non-sense AI generated content?
|
| Wasn't this always the outcome of the post labor economy?
|
| For this discussion lets just say that AI+Robots could replace
| most human labor and thinking. What do people do? Entertainment
| is going to be the number one time consumer.
| quantumHazer wrote:
| > just admit it, interns are struggling competing with Claude
| Code, Cursor and Codex
|
| They are not. This is false, zirp ended, this is the problem.
| Not LLMs.
| bopbopbop7 wrote:
| Try to provide some evidence first that AI is replacing people
| and that interns are struggling to compete with an LLM.
| ElijahLynn wrote:
| "download the Sora app"
|
| click
|
| takes me to the iPhone app store...
| m3kw9 wrote:
| I'm eagerly awaiting for some unexpected social problems this
| crops up
| sudohalt wrote:
| Now videos will be generated on the fly based on your preference.
| You will never put your phone down, it will detect when your sad
| or happy and generate videos accordingly
| intended wrote:
| That dragon flew backwards at one point didnt it.
|
| Impressive that THAT was one of the issues to find, given where
| we were at the start of the year.
| adidoit wrote:
| Impressive tech. Don't love the likely societal implications.
| joshdavham wrote:
| Will something like Sora 2 actually be used in Hollywood
| productions? If so, what types of scenes?
|
| I imagine it won't necessarily be used in long scenes with subtle
| body language, etc involved. But maybe it'll be used in other
| types of scenes?
| gamegoblin wrote:
| I saw a famous actor-director (can't remember who, but an
| A-list guy) said it would be super valuable even if you only
| use it for establishing shots.
|
| Like you have an exterior shot of a cabin, the surrounding
| environment, etc -- all generated. Then you jump inside which
| can be shot on a traditional set in a studio.
|
| Getting that establishing shot in real life might cost $30K to
| find a location, get the crew there, etc. Huge boon to indie
| films on a budget, but being able to endlessly tweak the shot
| is valuable even for productions that could afford to do it
| IRL.
| esafak wrote:
| Probably Ben Affleck.
| https://www.youtube.com/watch?v=ypURoMU3P3U
| deelowe wrote:
| Wow. What an intelligent take. I would have never expected
| this from Ben Affleck. He seems extremely familiar with the
| technology and it's capabilities and limits.
| gamegoblin wrote:
| Searched around and found it. It was actually Ashton
| Kutcher's interview with Eric Schmidt.
|
| Kutcher mentions the establishing shots, and I'd forgotten
| also points out the utility for relatively short stunt
| sequences.
|
| > Why would you go out and shoot an establishing shot of a
| house in a television show when you could just create the
| establishing shot for $100? To go out and shoot it would
| cost you thousands of dollars.
|
| > Action scenes of me jumping off of this building, you
| don't have to have a stunt person go do it, you could just
| go do it [with AI].
| echelon wrote:
| Casey Affleck is currently shooting a horror vampire period
| piece using Comfy UI and an Unreal Engine Volume. The AI is
| used for the background plates. It's just a test, but it's
| happening right now.
|
| Jason Blum is also getting really into the tech.
| plastic3169 wrote:
| People use these for sure but the biggest problem with these I
| feel is that they produce "finished shots" with 8 bit colors
| and heavy grading. It's hard to mix it with the other material
| which actually looks quite bland while it is being worked on.
| Would be great if somebody would train a model on raw footage.
| basisword wrote:
| Tens of billions in funding and they've just built a modern
| version of JibJab[1]. Can't wait to start receiving this in
| reply-all family emails.
|
| [1] https://youtu.be/z8Q-sRdV7SY?si=NjuyzL1zzq6IWPAe
| simonw wrote:
| The main lesson I learned from the March ChatGPT image generation
| launch - which signed up 100 million new users in the first week
| - is that people _love_ being able to generate images of their
| friends and family (and pets).
|
| I expect the "cameo" feature is an attempt at capturing that
| viral magic a second time.
| minimaxir wrote:
| Fortunately, you don't need permission from pets to use them in
| an AI video. (unless PETA objects)
| colonial wrote:
| Cool - now let's see how much it costs in compute to generate a
| single clip. (Also, notice how no individual scene is longer than
| a handful of seconds?)
| bergheim wrote:
| We are just heading for Lovely All TM.
|
| I kid.
|
| Art should require effort. And by that I mean effort on the part
| of the artist. Not environmental damage. I am SO tired of non
| tech friends SWOONING me with some song they made in 0.3 seconds.
| I tell them, sarcastically, that I am indeed very impressed with
| their endeavors.
|
| I know many people will disagree with me here, but I would be
| _heart broken_ if it turned out someone like Nick Cave was AI
| generated.
|
| And of course this goes into a philosophical debate. What does it
| matter if it was generated by AI?
|
| And that's where we are heading. But for me I feel effort is
| required, where we are going means close to 0 effort required.
| Someone here said that just raises the bar for good movies. I say
| that mostly means we will get 1 billion movies. Most are "free"
| to produce and displaces the 0.0001% human made/good stuff. I
| dunno. Whoever had the PR machine on point got the blockbuster.
| Not weird, since the studio tried 300 000 000 of them at the same
| time.
|
| Who the fuck wants that?
|
| I feel like that ship in Wall-E. Let's invest in slurpies.
|
| Anyway; AI is here and all of that, we are all embracing it. Will
| be interesting to see how all this ends once the fallout lands.
|
| Sorry for a comment that feels all over the place; on the tram :)
| GuinansEyebrows wrote:
| "if it's not worth [writing/playing/painting...], it's not
| worth [reading/listening/looking...]"
| bergheim wrote:
| I had a friend over for my last birthday before going to a
| venue. He had a huge framed painting he had made. It made me
| cry.
|
| A prompt delivered by Amazon drones would obviously not be
| the same lovely moment.
|
| So yes, I agree.
| IncreasePosts wrote:
| It's fitting that they host the video on Youtube, since that is
| where all of their training data came from.
| stan_kirdey wrote:
| That could totally power next generation of green-screen techs.
| Generative actors may not find favorable response in the
| audiences; but SFX, decor, extras, environments that react to
| actors' actions - amazing potential.
| portaouflop wrote:
| You can already do really cool stuff in this area "old" tech
| like stable diffusion. Not realistic or anything but really
| cool looking/morphing images
| adventured wrote:
| At least in terms of realism, the image generation field is
| at the realism line now. Single frame generation with Wan 2.1
| / 2.2 (and others) for example, will get you realism.
| zarzavat wrote:
| I can see that future generations are going to think that I'm
| boomer for preferring the performances of real actors instead
| of AI slop.
|
| The music industry already went through this with AutoTune and
| we know how that turned out.
| poisonarena wrote:
| >The music industry already went through this with AutoTune
| and we know how that turned out.
|
| they use it, everyone uses it, it got better to the point
| where most people dont know its used, ever heard of melodyne?
| well AI made it even better.
|
| And then there has been about 20 years of people using it
| even as their style of music, notably in hip hop, reggaeton,
| urbano, country, etc.
|
| Boomers like to think it was just an annoying fad in
| 2008-2011 or something, but it never went away, now everyone
| uses it, whether obvious or not
| r_lee wrote:
| I don't get the autotune argument. It's like saying we
| shouldn't be using electronic instruments because it's not
| real or we shouldn't use digital audio instruments because
| they're not real etc.
|
| It's just a way to get different kind of sound. It won't make
| you good tracks.
| Rudybega wrote:
| I think AI is starting to verge on making actual good
| music. The latest Suno release is wild.
|
| An example here: https://v.redd.it/fqlqrgumo5rf1
|
| I find this one interesting because Rap has classically
| been difficult for these models (I think because it's
| technically difficult to find the right rhythms and flow
| for a given set of lyrics).
| kfajdsl wrote:
| > The music industry already went through this with AutoTune
| and we know how that turned out.
|
| Yeah, it turned out that almost all mainstream tracks
| nowadays have post-processing on vocals (the extent varying
| between genres and styles).
| Tiktaalik wrote:
| There's less here than you think. Video games have already been
| procedurally generating environment art for quite some time,
| and film/tv are already leveraging that with giant screens that
| use Unreal Engine to create the backgrounds.
|
| AI could be helpful here, but it's not clear that it is
| required or an improvement.
| thebiglebrewski wrote:
| Can this be used to make hyper-realistic video games, or it's not
| that real-time yet?
| dagaci wrote:
| Amazing. iOS only, with region restrictions in 2025.
| asadm wrote:
| considering legal foolishness of EU, this is the right move.
| TheAceOfHearts wrote:
| > Sora is not available in Puerto Rico yet
|
| I love the casual reminds that we're second-class citizens each
| time a new technology gets released. Available in the US but
| always excluding Puerto Rico.
| GaggiX wrote:
| The model's quality is incredible, but more tools are needed to
| take advantage of its capabilities, this is kinda the magic of
| open models.
| barbarr wrote:
| Instagram reels are gonna get crazy
| artursapek wrote:
| You see the one with the dolphin on the trampoline?
| ashu1461 wrote:
| Those `nature is amazing type of videos` are already flooded
| with AI
| MangoToupe wrote:
| Interesting that they're going with a "copyright opt-out":
| https://www.reuters.com/technology/openais-new-sora-video-ge...
|
| I guess copyright is pretty much dead now that the economy relies
| on violating it. Too bad those of us not invested into AI still
| won't be able to freely trade data as we please....
| alkonaut wrote:
| How far out are we from doing this in real time? What's the
| processing/rendering time per frame?
| kachapopopow wrote:
| could already do it in real time by dimming the lightbulbs of a
| city or two.
| neom wrote:
| https://deepmind.google/discover/blog/genie-3-a-new-frontier...
| beders wrote:
| Can I finally redo the Star Wars sequels with this? :)
| crims0n wrote:
| Didn't Star Wars end in 2005?
| d--b wrote:
| Ok that's technically really impressive, and probably totally
| unusable in a real creativity context beyond stupid ads and
| politically-motivated deepfakes.
| deng wrote:
| As usual: impressive until you look close. Just freeze the frame
| and you see all the typical slop errors: pretty much any kind of
| writing is a garbled mess (look at the camera in the beginning).
| The horn of the unicorn sits on the bridle. The buttons on Sam's
| circus uniform hover in the air. There are candleholders with
| somehow candles inside as well as on top. The miniature
| instruments often make no sense. The conductor has 4 fingers on
| one hand and 5 on the other. The cheers of the audience is
| basically brown noise. Nedless to say, if you freeze the
| audience, hands are literally all over the place. Of course,
| everything conveniently has a ton of motion blur so you cannot
| see any detail.
|
| I know, I know. Most people don't care. How exciting.
| rendleflag wrote:
| Is your complaint that it has errors? I mean look at what it
| can do. This is a freaking computer generating things from
| scratch based on a prompt. Two years ago, technology like this
| was so much worse and could only generate basic images and
| videos. Now it can generate visuals all from the text someone
| puts in.
|
| Anyone, literally anyone, can use it (eventually) to generate
| incredible scenes. Imagine the person who comes up with a short
| film about an epic battle between griffins and aliens...Or a
| simple story of a boy walking in the woods with their dog...Or
| a story of a first kiss. Previously people were limited to what
| they had at hand. They couldn't produce a video because it was
| too costly. Now they can craft a video to meet their vision.
|
| I do find it exciting.
| deng wrote:
| > Is your complaint that it has errors?
|
| Well, yes? There's a reason why everything that was produced
| with these tools so far is garbage: because no one actually
| caring about their art would accept these things. Art is a
| deliberate thing, it takes effort. These tools are fine for
| company training videos and TikToks. Of course a few years
| ago this was science fiction. They are immensely impressive
| from a technical perspective. Two things can be true.
| bopbopbop7 wrote:
| There is that magic word again, "eventually". When is that?
| The same time we get warp drives?
| ascorbic wrote:
| This is super cool and fun and will almost certainly be really
| bad for society in loads of different ways. From the descriptions
| of all the guardrails they're needing to put in it seems like
| they know it too.
| bbor wrote:
| Glad to see someone is looking out for a forest, here. A
| diverse host of excuses have cropped up to explain away the
| anxiety AGI brings, and I totally understand why. Yet again,
| today we stare into the abyss. Sora 2
| represents significant progress towards [AGI]. In keeping with
| OpenAI's mission, it is important that humanity benefits from
| these models as they are developed.
|
| This seems like a good time to remind ourselves of the original
| OpenAI charter:
| https://web.archive.org/web/20230714043611/https://openai.co...
|
| I wonder how exactly they reconcile the quote above with "We
| are concerned about late-stage AGI development becoming a
| competitive race without time for adequate safety
| precautions"...
| nurettin wrote:
| I am not for or against AGI, but why is there anxiety around
| it? Do people simply hear sales rhetoric and assume that it
| can exist and will be used in order to dominate their lives?
| askl wrote:
| But think of all the 0 legitimate use cases for this
| technology.
| giancarlostoro wrote:
| So being able to generate sign language videos for people who
| cannot hear is not a legitimate use case for AI videos? Or is
| your hate boner for AI just blinding you from useful
| applications?
| umanwizard wrote:
| Why can't sign language be written? Why does it need to be
| on video?
| ascorbic wrote:
| There isn't a standard written form of any major sign
| languages
| 542458 wrote:
| Is there a reason that's superior to subtitles, which are
| already fairly easy to generate?
| currymj wrote:
| sign languages are completely different languages from
| spoken languages, with their own grammar etc.
|
| subtitles can work but it's basically a second language.
| perhaps comparable to many countries where people speak a
| dialect that's very different from the "standard" written
| language.
|
| this is why you sometimes have sign language interpreters
| at events, rather than just captions.
|
| there's not really a widely accepted written form of sign
| language.
| fluoridation wrote:
| >this is why you sometimes have sign language
| interpreters at events, rather than just captions.
|
| No, the reason is because a) it's in real time, and b)
| there's no screen to put the subtitles on. If it was
| possible to simply display subtitles on people's vision,
| that would be much more preferable, because writing is a
| form of communication more people are familiar with than
| sign language. For example, someone might not be deaf,
| but might still not be able to hear the audio, so a sign
| language interpreter would not help them at all, while
| closed captions would.
| currymj wrote:
| if you're maximizing accessibility you'd have both. often
| in broadcasts with closed captioning, there will also be
| a video of the sign language interpreter.
| fluoridation wrote:
| LOL. Yeah, that's way better than closed captions, even
| auto-generated ones.
| overfeed wrote:
| Holy over-engineering batman! Is text too old-fashioned?
| margalabargala wrote:
| Great point. Really, the main problem with subtitles is
| that the creator can understand them without having to know
| another language, and therefore can spot check them. That
| makes it much more difficult to insert Black Mirror-style
| Contextually Relevant Advertisements.
| dang wrote:
| Please don't respond to a bad comment by breaking the site
| guidelines yourself. That only makes things worse.
|
| (Your comment would be just fine without the last sentence)
|
| https://news.ycombinator.com/newsguidelines.html
| tkamado wrote:
| it helps altman with world domination, so one legitimate use
| case for one person?
| dang wrote:
| " _Please don 't post shallow dismissals, especially of other
| people's work. A good critical comment teaches us
| something._"
|
| " _Don 't be snarky._"
|
| https://news.ycombinator.com/newsguidelines.html
| minimaxir wrote:
| tbh I didn't know it was technically possible for dang to
| be downvoted
| haolez wrote:
| One use that occurred to me is that fans will be able to "fix"
| some movies that dropped the ball.
|
| For example, I saw a lot of people criticizing "Wish" (2023,
| Disney) for being a good movie in the first half, and totally
| dropping the ball in the last half. I haven't seen it yet, but
| I'm wondering if fans will be able to evolve the source material
| in the future to get the best possible version of it.
|
| Maybe we will even get a good closure for Lost (2004)!
|
| (I'm ignoring copyright aspects, of course, because those are too
| boring :D)
| BeetleB wrote:
| Or just going to the Goofs section of a movie on IMDB, and fix
| the trivial issues (e.g. car had cracked window in earlier
| scene, and suddenly a normal window in another scene).
|
| Much more mundane, but useful!
| ronsor wrote:
| > (I'm ignoring copyright aspects, of course, because those are
| too boring :D)
|
| You must understand that infinite copyright is the author's
| right, and AI companies must be sued for 50 trillion dollars.
| haolez wrote:
| Come on. This is just a fun thought exercise. I'm not
| suggesting creating a startup around this.
| ronsor wrote:
| I was trying my hand at satire; but I understand that many
| people now genuinely hold such extreme views.
| SkyBelow wrote:
| My issue is that the copyright aspect are what prevents me from
| using this as much as I otherwise would.
|
| About 6 months ago I asked a few different AIs if they could
| translate a song for me as a learning experience, meaning not a
| simple translation, but more a word by word explanation of what
| each word meant, how it was conjugated, any more
| musical/lyrical only uses that aren't common outside of songs,
| and so on. I was consistently refused on copyright grounds,
| despite this seeming a fair use given the educational nature.
| If I pasted a line of the lyrics at a time, it would work
| initially, but eventually I would need to start a new chat
| because the AI determined I translated too much at once.
|
| So in this one, if I wanted to ask it to create a video of the
| moment in Final Fantasy 6 when the bad guy wins, or a video of
| the main characters of Final Fantasy 7 and 8 having a sword
| duel, would it outright refuse for copyright reasons?
|
| It sounds like it would block me, which makes me lose a bit of
| interest in the technology. I could try to get around it, but
| at what point might that lead to my account being flagged as a
| trouble maker trying to bypass 'safety' features. I'm hoping in
| a few years the copyright fights on AI dies down and we get
| more fair use allowance instead of the tighter limitations to
| try to prevent calls for tighter regulation.
| inerte wrote:
| Just yesterday I learned "This summer, two Dramione fics turned
| rewritten novels became New York Times bestsellers" -
| https://slate.com/culture/2025/09/alchemised-senlinyu-harry-...
|
| 100% sure we will see people re-doing movie parts. Also see
| https://en.wikipedia.org/wiki/The_Phantom_Edit
| qgin wrote:
| VFX artists are definitely feeling the AGI / considering other
| career paths today.
| Banditoz wrote:
| I genuinely don't understand the consistent rhetoric on this
| site of:
|
| > new AI feature/model comes out
|
| > "it's going to replace people in this field! they better
| start looking for a new job!!!"
|
| why is this a good thing?
| qgin wrote:
| It's not a good thing, but it's definitely a thing. Most of
| us here on HN are going to be affected by this.
| vultour wrote:
| How many more years do you think you'll need to keep saying
| this before it's actually true?
| Retr0id wrote:
| Who said it was a good thing?
| bopbopbop7 wrote:
| Is this AGI in the room with us now?
| myahio wrote:
| Not with the way this thing renders hair (or any other high
| fidelity texture)
| https://x.com/GabrielPeterss4/status/1973090475486879818
| bsenftner wrote:
| VFX artist and developer here, who's deep into this stuff, and
| it is really not there. It's an island of itself, barely
| controllable and barely usable with other media. They are just
| now getting around to generating alpha channels, with virtual
| none of the existing pipelines for any AI video or image
| generation tools to even incorporate and work with alpha
| channels. This is just one of several hundred aspects of
| incompatibility. It really seriously appears as of no one at
| any of the AI video generation research teams has any
| professional media production experience, or even bothered too
| look at existing media production data standards, and what they
| are making tool-wise is incompatible in every possible respect.
| rhetocj23 wrote:
| "It really seriously appears as of no one at any of the AI
| video generation research teams has any professional media
| production experience, or even bothered too look at existing
| media production data standards,"
|
| I had to chuckle at this. Because the arrogance of OAI et al
| will finally get them in the end when these projects continue
| to be negative NPV.
| gmueckl wrote:
| Do you even see a path from the current AI systems to
| something that has that near-total control over every detail
| that is required for high quality VFX work?
| dragonwriter wrote:
| "With Sora 2, we are jumping straight to what we think may be the
| GPT-3.5 moment for video."
|
| I think feeling like you need to use that in marketing copy is a
| pretty good clue in itself both that its not, and that you don't
| believe it is so much as desperately wish it would be.
| echelon wrote:
| The Sora app squaring off against Meta's social video app is
| the real story here.
|
| Sora 2 itself looks and sounds a little poorer than Google Veo
| 3. (Which is itself not currently ranked as the top video
| model. The Chinese models are dominating.)
|
| I think Google, with their massive YouTube data set, is
| ultimately going to win this game. They have all the data and
| infrastructure in the world to build best-in-class video
| models, and they're just getting started.
|
| The social battle will be something completely different,
| though. And that's something that I think OpenAI stands a good
| chance at winning.
|
| Edit: Most companies that are confident of their image or video
| models stealthily launch it on the Model Arena a week ahead of
| the public model release. OpenAI did not arrange to do that for
| Sora 2.
|
| Nano Banana, Seedream/Seedance, Kling, and several other models
| have followed this pattern of "stealth ELO ranking, then reveal
| pole position".
|
| https://artificialanalysis.ai/text-to-video/arena?tab=leader...
|
| The fact that this model is about "friends" and "social"
| implies that this is an underpowered model. You probably saw a
| cherry picked highlight reel with a large VRAM context, but the
| actual consumer product will be engineered for efficiency.
| Built to sustain a high volume of cheap generations, not
| expensive high quality ones. A product built to face off
| against Meta. That model compete on the basis of putting you
| into videos with Pikachu, Mario, and Goku.
| CaptainOfCoit wrote:
| > I think Google, with their massive YouTube data set, is
| ultimately going to win this game.
|
| I don't know, applying the same thinking to LLMs, Google
| should have been first and best with just text based LLMs
| too, considering the datasets they sit on (and researchers,
| among others the people who came up with attention). But
| OpenAI somehow beat them on that regardless.
| jstummbillig wrote:
| I am looking at the videos and really had a feeling that it
| looks right (minus a lot of obvious fuck ups still) where
| previously something felt fundamentally wrong with ai videos.
| It feels _somewhat_ important, in so far you consider ai
| generated videos important.
| horhay wrote:
| It's the skin textures. It's the slightly better lipsyncing.
| Maybe it will be different when us normal users get it but so
| far the demos with Sam don't make him look waxy.
| gainda wrote:
| impressive engineering that's hard to see as a net good for
| humanity.
|
| it doesn't spark optimism or joy about the future of engaging
| with the internet & content which was already at a low point.
|
| old is gold, even more so
| dyauspitr wrote:
| How did they generate the videos with Sam Altman. Did they just
| provide a picture of his face and then use him in their prompts?
| rodonn wrote:
| You can use the "cameo" feature only with users who have gone
| through the cameo creation flow. Sama has an account and
| created a cameo likeness of himself. When you create your cameo
| you can choose who is allowed to make videos using it: "only
| me", "people I approve", "mutuals", or "everyone".
| kaicianflone wrote:
| Why is the video player so laggy?
| cubefox wrote:
| Right? It constantly dropped frames for me (Firefox/Android).
| darkwater wrote:
| Last famous words:
|
| > A lot of problems with other apps stem from the monetization
| model incentivizing decisions that are at odds with user
| wellbeing. Transparently, our only current plan is to eventually
| give users the option to pay some amount to generate an extra
| video if there's too much demand relative to available compute.
| As the app evolves, we will openly communicate any changes in our
| approach here, while continuing to keep user wellbeing as our
| main goal.
| Workaccount2 wrote:
| Sam will quickly learn that general users give -zero- thought
| to OpenAI well being. Nor be bothered that they should give it
| a thought.
| ambicapter wrote:
| AI Sam Altman is terrifying, holy shit. Squarely in uncanny
| valley for me.
| benzible wrote:
| Came here to say this myself. Would like to unsee that.
| neom wrote:
| Going to be an amazing source of training data, wait till they
| get it to real time and people are leaving their video camera
| open for AR features. OpenAI is about to have a lot of current
| real world image data, never mind the sentiment analysis.
| altcognito wrote:
| I don't think they were limited for video training data.
| Gathering real world data is pretty easy, gathering curated
| information is a little more difficult.
| bovermyer wrote:
| "Thou shalt not create a machine in the likeness of a human
| mind."
| saguntum wrote:
| I wonder if they're going to license this to brands for heavily
| personalized advertisement. Imagine being able to see videos of
| yourself wearing clothes you're buying online before you actually
| place the order, instead of viewing them on a model.
|
| If they got the generation "live" enough, imagine walking past a
| mirror in a department store and seeing yourself in different
| clothes.
|
| Wild times.
| foota wrote:
| The latter would feel like actual scifi to me.
| larodi wrote:
| its called Virtual Try On (VTO) and there are plenty of models
| going there for static gfx, it is very reasonable to expect
| soon emerge those for video VTO.
| shubb wrote:
| Accurate virtual try on however is quite difficult, and users
| will quickly learn to distrust platforms that just generate
| something that"looks right".
|
| You can prompt with a normal size 8 dress and "kim jungle un
| wearing a dress" and it will show you something that doesn't
| help you understand whether that dress would fit or not. You
| can ask for a tube dress and it will usually give him a big
| bust to hold it up. It's not useful for the purpose of
| visualing fit.
|
| It will definitely be used for such just like image models
| already are for cheap tenu clothes, and our onions shopping
| experience will get worse.
|
| Maybe this needs purpose built models like vibe-net or maybe
| you cab train a general purpose model to do it, but if they
| were spending the effort necessary to do so they'd be calling
| it out.
| cyrialize wrote:
| I'm fairly certain there is a scene in Minority Report just
| like this! Or at least, the advertisement says Tom Cruise's
| character's name.
|
| https://en.wikipedia.org/wiki/Minority_Report_(film)
| beklein wrote:
| Here a clip of that scene:
| https://www.youtube.com/watch?v=7bXJ_obaiYQ
| echelon wrote:
| In 2023, Carvana ran an ad campaign that showed you a video
| of "your car" thanking you and talking about your time
| together:
|
| https://adage.com/article/digital-marketing-ad-tech-
| news/car...
|
| A little creepy, but very much in this vein.
|
| We probably haven't even scratched the surface of what will
| be done with this tech. When video becomes "easy", "quick",
| "affordable", and "automatable" (something never before
| possible on any of those dimensions) - it enables countless
| new things to be done.
| iLoveOncall wrote:
| You don't need generative AI for that at all, snapchat filters
| have existed for a decade and are the same concept. A lot of
| brands have already adopted that.
| gm678 wrote:
| Or, on the genAI side, Google marketed this use case heavily
| for Flash Image 2.5 (even if that's not the same type of
| generative model because it's geared for editing, it's still
| in the taxonomy)
| seydor wrote:
| When the dust settles , that's probably going to be the most
| common application of these video models. Making automated
| social content kind of defeats the purpose; people empathize
| with other people, not with AI . (I guess that's why they
| didn't also make their interview video via AI)
|
| But Sora /VEO will probably also revolutionize movies and tv
| content
| busymom0 wrote:
| Am I misremembering or didn't Meta announce few months ago that
| people will see their own faces in ads?
| latexr wrote:
| At that point, why even buy the clothes? Influencers will just
| post the video of the mockup on social media, which is the only
| reason they were considering it in the first place. Save
| themselves the foot fungus.
|
| https://xcancel.com/Naija_PR/status/1904809073356251634
|
| Then take the next step. Why even spend money going out?
| Generate a video of yourself with fake friends at a party and
| post that, while eating ice cream alone at home.
| ares623 wrote:
| Now you're thinking with portals
| chilipepperhott wrote:
| People said the exact same thing about AR furniture, and I'm
| 99% sure no one uses that.
| dwa3592 wrote:
| I don't know if it's just me or other people are feeling it as
| well. I don't enjoy videos anymore (unless live sports). I don't
| enjoy reading on my monitor anymore, I have been going back to
| physical books more often. I am in my early thirties.
|
| The point is that sora2 demo videos seemed impressive but I just
| didn't feel any real excitement. I am not sure who this is really
| helping.
| marcofloriano wrote:
| Same with me !
| greenavocado wrote:
| Personally I can't wait for super creative and novel indie film
| productions as film production will be more liberated from the
| grip of Hollywood and the influence of the upper classes in
| general. Especially once the Chinese make less-censored-to-
| Western-users models more available and even more so once
| people can run these things at home in some years.
| kobalsky wrote:
| that sounds like clinical depresion, I'd check with my
| endocrinologist to get blood work done
| marcofloriano wrote:
| Every AI video demonstration is always about funny stuff and
| fancy situations. We never see videos on art, history,
| literature, poetry, religion (imagine building a video about the
| moment Jesus was born) ... ducks in a race !? Come on ...
|
| So much visual power, yet so little soul power. We are dying.
| fluoridation wrote:
| What do you imagine a generated video about poetry would be?
|
| >Every AI video demonstration is always about funny stuff and
| fancy situations.
|
| The thing about AI slop is that by its very nature, unless it's
| heavily reined in by a human, it's invariably lowest common
| denominator garbage. It very likely will generate something you
| yourself could think of within the first five seconds of
| hearing the prompt, not some very clever take on it, so it can
| _only_ work as a placeholder (AI as a replacement of stock
| images is great, for example) or to add background detail where
| it won 't call attention to itself and its genericity.
|
| >imagine building a video about the moment Jesus was born
|
| Given there are multiple paintings on the subject, I very much
| doubt no one has generated something like that already.
| boh wrote:
| This is the kind of thing people get excited about for the first
| couple of months and then barely use it going forward. It's
| amazing how quickly the novelty of this amazing technology wears
| off. You realize how necessary meaning/identity/narrative is to
| media and how empty it gets (regardless of the output) when those
| elements are missing.
| tptacek wrote:
| If I was on the OpenAI marketing team I maybe wouldn't have
| included the phrase "and letting your friends cast you in their
| [videos]". It's a little chilling.
| minimaxir wrote:
| The livestream showed an interesting UX with Facebook-style
| permissions that make it so you very explicitly have to opt
| into this feature:
| https://bsky.app/profile/minimaxir.bsky.social/post/3m22zg2h...
|
| Even moreso than Facebook tags, the person being cast can cause
| the deletion of the source video at any time.
| drcongo wrote:
| The AI generated Sam Altman doesn't look even vaguely human.
| echelon wrote:
| I'm a software engineer and hobbyist actor/director. My friends
| are in the film industry and are in IATSE and SAG-AFTRA. I've
| made photons-on-glass films for decades, and I frequently film
| stuff with my friends for festivals.
|
| I love this AI video technology.
|
| Here are some of the films my friends and I have been making with
| AI. These are not "prompted", but instead use a lot of hand
| animation, rotoscoping, and human voice acting in addition to AI
| assistance:
|
| https://www.youtube.com/watch?v=H4NFXGMuwpY
|
| https://www.youtube.com/watch?v=tAAiiKteM-U
|
| https://www.youtube.com/watch?v=7x7IZkHiGD8
|
| https://www.youtube.com/watch?v=Tii9uF0nAx4
|
| Here are films from other industry folks. One of them writes for
| a TV show you probably watch:
|
| https://www.youtube.com/watch?v=FAQWRBCt_5E
|
| https://www.youtube.com/watch?v=t_SgA6ymPuc
|
| https://www.youtube.com/watch?v=OCZC6XmEmK0
|
| I see several incredibly good things happening with this tech:
|
| - More people being able to visually articulate themselves,
| including "lay" people who typically do not use editing software.
|
| - Creative talent at the bottom rungs being able to reach high
| with their ambition and pitch grand ideas. With enough effort,
| they don't even need studio capital anymore. (Think about the
| tens of thousands of students that go to film school that never
| get to direct their dream film. That was a lot of us!)
|
| - Smaller studios can start to compete with big studios. A ten
| person studio in France can now make a well-crafted animation
| that has more heart and soul than recent by-the-formula Pixar
| films. It's going to start looking like indie games. Silksong and
| Undertale and Stardew Valley, but for movies, shows, and shorts.
| Makoto Shinkai did this once by himself with "Voices of a Distant
| Star", but it hasn't been oft repeated. Now that is becoming
| possible.
|
| You can't just "prompt" this stuff. It takes work. (Each of the
| shorts above took days of effort - something you probably
| wouldn't know unless you're in the trenches trying to use the
| tech!)
|
| For people that know how to do a little VFX and editing, and that
| know the basic rules of storytelling, these tools are remarkable
| assets that compliment an existing skill set. But every shot,
| every location, every scene is still work. And you have to weave
| that all into a compelling story with good hooks and visuals.
| It's multi-layered and complex. Not unlike code.
|
| And another code analogy: think of these models like Claude Code
| for the creative. An exoskeleton, but not the core driving
| engineer or vision that draws it all together. You can't prompt a
| code base, and similarly, you can't prompt a movie. At least not
| anytime soon.
| Mashimo wrote:
| Well I was entertained.
|
| What is up with a lot of voices are left ear only?
| echelon wrote:
| Carter needs a new laptop. His daily driver has been falling
| apart for ages but he refuses to give it up.
|
| We all told him about the sound mix - he let a couple of
| videos slip with a bad "mono as single-channel stereo audio"
| renders. On his machine it sounded normal. He got flack for
| that, and he's been hearing this for months.
|
| I'm going to show him this thread. I don't think he'll ever
| forget to check again.
|
| Despite that, he's a really talented guy. Chalk this up as a
| bad production deploy. We didn't want to delete and re-upload
| since the videos had legs when we first released them.
| There's a checklist now.
| summarity wrote:
| Lol I wish YT had a warning for that.
|
| In the meantime, good old
|
| Settings -> Accessibility -> Audio -> Play Stereo as Mono
|
| helps.
| marseysneed wrote:
| My left ear enjoyed these videos
| tobr wrote:
| Adding to the list: The Adventures of Reemo Green. Very funny,
| and the first time I've watched AI video and enjoyed it as more
| than a technical curiosity.
|
| https://www.youtube.com/watch?v=5bYA2Rv2CQ8
| kingds wrote:
| sorry but it's funny that you mention "heart and soul" while
| sharing some of the most soulless videos i've ever seen.
| echelon wrote:
| I'll have you know that in this year's Atlanta 48 Hour Film
| project (something I've been doing since I was a teen),
| several teams used AI.
|
| Rewind to just one year prior -- 2024.
|
| AI video was brand-spanking new. We'd only just gotten over
| the "Will Smith" spaghetti video and the meme-y "Pepperoni
| Hug Spot" and "Harry Potter by Balenciaga" videos.
|
| I was the only person to attempt to use AI in 2024's
| competition. It was a time when the tools and infrastructure
| for video barely existed.
|
| On the debut night, I was resoundingly booed by the audience.
| It felt surreal. Working all weekend to have an audience of
| peers jeering at you in a dark theater. The judges gave me an
| award out of sympathy.
|
| Back then, image-to-video models really were not a thing
| (Luma launched "Dream Machine v1" shortly after this). I was
| using Comfy, Blender, Mocap, a full Mocap suit (the itchy
| kind), and a lot of other hacks to build something with
| extremely crude tools.
|
| We lost a day of filming and had to scramble to get something
| done in just 24 hours. No sleep, too much caffeine. Lots of
| sweat and toil.
|
| The resulting film was a total mess, of course:
|
| https://vimeo.com/955680517/05d9fb0c4f (It's seriously bad -
| I hate it. It might legitimately be the very first time AI
| was used in a 48 hour competition.)
|
| That said, it felt very much like a real 48 Hour competition
| to me. Like a game jam. The crude ingredients, an idea, the
| clock. The hustle. The corners being cut. It was palpable.
|
| I don't think you can say there isn't soul in this process.
| The process has so much soul.
|
| Anyway, fast forward to this year. Three teams used AI,
| including my own. (I don't think I have a link to our film,
| sadly.)
|
| We all got applause. The audience was full of industry folks,
| students, and hobbyists. They loved it. And they knew we used
| AI.
|
| The industry is anxious but curious about the tech. But
| fundamentally, it's a new tool for the tool box. The real
| task is storytelling.
| mintone wrote:
| I wrote this a year or so ago:
| https://www.technicalchops.com/articles/ai-goes-to-hollywood...
|
| "The studios and creators who thrive in this new landscape will
| be those who can effectively harness AI's capabilities while
| maintaining the human creativity and vision that ultimately
| drives the art of cinema."
|
| It is in many ways thrilling to see this come to life, and I
| couldn't agree with you more.
| hansmayer wrote:
| > "The studios and creators who thrive in this new landscape
| will be those who can effectively harness AI's capabilities
| while maintaining the human creativity and vision that
| ultimately drives the art of cinema."
|
| ..Just somehow several years on, these optimistic statements
| still all end up being in the future tense, somehow for all
| the supposed greatness and benefits, we still dont see really
| valuable outputs. A lot of us do not want more of the
| "CONTENT" as envisioned by corporate ghouls who want their
| employees or artists to "thrive" (another word kidnapped by
| LinkedIn-Linguists). The point is not in the speed and
| easiness of generation of outputs, visual and sound effects
| etc. The point is the artists interpretation and their own
| vision, impressions etc. Not a statistical slop which
| "likely" fits my preferences (i.e. increases my dopamin
| levels).
| squidsoup wrote:
| Creative people with ambition and limited resources make good
| things today without this technology. All this does is
| accelerate the rate at which low quality "content" is produced
| by people that have no interest in learning a craft, without
| attribution and without compensation for the people that have
| made the effort and whose works train these models.
| rhetocj23 wrote:
| Precisely.
|
| I have a really big problem with letting low quality stuff
| infest into the species.
| bnop wrote:
| Taking the time and effort out of something is exactly what
| strips it of its beauty
|
| Beauty is not just an "idea" that someone has and needs to get
| out onto a medium
|
| It is a process and journey that a person undergoes to get said
| idea onto said medium
|
| That journey often plays out very differently than the person
| expects. Things change, the art is different from the idea, and
| the person learns and grows
|
| Our modern society is so obsessed with results, competition,
| and efficiency that we no longer see the truth: the journey is
| to be enjoyed, and from enjoying the journey, comes beauty
|
| I encourage you to meditate on why our society is so sick and
| depressed right now, and extrapolate to how we got here, before
| assuming this will be a good thing for society
| sumeruchat wrote:
| Shameless plug but I am creating a startup in this space called
| cleanvideo.cc to tackle some of the issues that will come with
| fake news videos. https://cleanvideo.cc
| robotsquidward wrote:
| It's insanely impressive. At the same time, all these videos all
| look terrible to me. Still get extreme uncanny valley and
| literally makes me sick to my stomach.
| spaceman_2020 wrote:
| This stuff works really well when you make something that's
| exaggerated reality, as in either an animation or a MTV-style
| music video
|
| I can't find the link now, but I saw a continuous shot video of
| a grocery store from the perspective of a fly. It was shot in
| the 90s music video style and looked so damn good.
|
| Some of the stuff being done by these guys is also a whole lot
| of fun (slightly NSFW and political content), and it fits the
| music video theme:
|
| https://www.youtube.com/watch?v=V4zwIhS2iZk
| jrop wrote:
| Agree - leaps and bounds beyond anything I would have dreamed
| possible a few years ago...but... IDK, if I'm honest, the sound
| was way off too, not just the visuals. The music sounded
| detuned slightly, and the crowd noise was "crackly" etc. etc.
| It had a low-fidelity "quality" to it.
|
| Personally, I feel mixed feelings. I'm impressed, but I'm not
| looking forward to the new "movies" that are going to litter
| YouTube et al generated from this.
| unsnap_biceps wrote:
| They seem like they're low FPS videos. I wonder if they're
| rendering 24 FPS and it's mismatching youtube's 30 FPS and
| causing the weird stuttering.
| unethical_ban wrote:
| I just had a thought: (spoilers Expanse and Hyperion and Fire
| Upon the Deep)
|
| Multiple sci-fi-fantasy tales have been written about technology
| getting so out of control, either through its own doing or by
| abuse by a malevolent controller, that society must sever itself
| from that technology very intentionally and permanently.
|
| I think the idea of AGI and transhumanism is that moment for
| society. I think it's hard to put the genie back in the bottle
| because multiple adversarial powers are racing to be more
| powerful than the rest, but maybe the best thing for society
| would be if every tensor chip disintegrated the moment they came
| into existence.
|
| I don't see how society is better when everyone can run their own
| gooner simulation and share it with videos made of their high
| school classmates. Or how we'll benefit from being unable to
| trust any photo or video we see without trusting who sends it to
| you, and even then doubting its veracity. Not being able to hear
| your spouse's voice on the phone without checking the post-
| quantum digital signature of their transmission for authenticity.
|
| Society is heading to a less stable, less certain moment than any
| point in its history, and it is happening within our lifetime.
| sys32768 wrote:
| I welcome a world where gullible people begin to doubt everything
| they see.
| kibwen wrote:
| My friend, no man ever got rich betting against the infinitely
| deep well of human gullibility.
| skybrian wrote:
| The result of that isn't rational skepticism, though. It's
| distrusting mainstream news and embracing whatever conspiracy
| theories your friends believe.
| polishdude20 wrote:
| There's something about the faces that looks completely off to
| me. I think it's the way the mouth and whole face moves when they
| talk.
| HarHarVeryFunny wrote:
| Yeah, the faces aren't right, and impressive as it is I'm
| getting icky "uncanny valley" vibes from this.
|
| CGI for fantasy stuff is unavoidable, but when it's stuff that
| could have been done by actors but is instead AI, then to me it
| just feels cheap and nasty - fake.
| bob1029 wrote:
| It's the inaccuracy of things like shadows, sub-surface
| scattering and specular highlights. I think the shadow
| inaccuracy is what the human visual system is most sensitive
| to.
|
| These LLMs might make content that looks initially impressive
| but they are absolutely not performing physically based
| rendering or have any awareness of the lighting arrangement in
| these scenes. There are a lot of things they get right, but you
| only have to screw up one small element to throw the whole
| thing off.
|
| I am willing to bet that Unreal Engine 5 will continue to
| produce more realistic human faces than OAI ever can with these
| types of models. You cannot beat the effects of actually
| running raytracing in a PBR pipeline.
| SeanAnderson wrote:
| Sheeeeeeeeeeesh. That was so impressive. I had to go back to the
| start and confirm it said "Everything you're about to see is Sora
| 2" when I saw Sam do that intro. I thought there was a prologue
| that was native film before getting to the generated content.
| iLoveOncall wrote:
| I'm sorry but that's a gross exageration. If any of this was
| real film then I'd start a gofundme page for OpenAI to get
| better video production equipment and team because that would
| be laughably bad.
|
| If anything, it looks a lot worse than a lot of AI-generated
| videos I've seen in the past, despite being a tech demo with
| carefully curated shots. Veo 3 just blows this out of the water
| for example.
| SeanAnderson wrote:
| It's not an exaggeration to me? I literally stopped the video
| and went back to the start and re-read. You're more than
| welcome to speak about your opinions and experiences, but I'm
| speaking about mine.
|
| I'm over here thinking, "It felt like just yesterday I was
| laughing at trippy, incoherent videos of Will Smith eating
| spaghetti."
|
| I love the progress we're making. I love the competition
| between big companies trying to make the most appealing
| product demos. I love not knowing what the tech world is
| going to look like in six months. I love not thinking, "Man.
| The Internet was a cool invention to have grown up in, but
| now all tech is mundane and extractive." Every time I see AI
| progress I'm filled with childlike wonder that I thought was
| gone for good.
|
| I don't know if this represent SOTA for video generation. I
| don't care. In that moment I found it impressive and was
| commenting specifically on the joy I experienced watching the
| video. I find it frustrating to have that joy met with such
| negativity.
| ryandrake wrote:
| Don't worry. AI is going to be monetized and extractive in
| no time. Just like Social Media went from "fresh, fun and
| cool new tech" to "how did we let this horrible beast take
| hold of the world," AI will take the same path. In 10 years
| or sooner, when 99.99% of what you read, hear, and watch is
| AI slop, you're going to post "This used to be a cool
| invention!" if there's even a place left for humans to post
| by that time.
| SeanAnderson wrote:
| I agree. It will absolutely get there. Such is the trend
| of all scientific inventions. A breakthroughs occurs,
| prosperity follows in response, hedonic adaption causes
| satisfaction to regress to the mean, and then people
| squeeze every remaining drop of value out of the
| technology while we wait for those capable of true
| innovation to work their magic once more. I don't find it
| idyllic, but I accept it as the way the world works. It
| feels like a force of nature to me.
|
| The period we're in is fleeting. I think it should be
| acknowledged and treasured for what it is rather than
| viewed with disdain because of what is inevitably to
| come. I stopped using Facebook and never moved to
| Insta/TikTok when things began to feel too extractive,
| but, for a good decade there, I felt so close to so many
| more people than I ever thought possible. It was a really
| nice experience that I no longer get to have. I'm not mad
| at social media. I'm happy I got to experience that
| window of time.
|
| Right now I'm very happy to be using LLMs without feeling
| like I'm being preyed upon. I love that programming feels
| fresh and new to me after 15 years. I'm looking forward
| to having my ability to self-express magnified ten-fold
| by leveraging generative audio/visuals, and I look
| forward to future breakthroughs that occur once all these
| inventions become glorified ad-delivery mechanisms.
|
| None of this seems bad to me. Innovation and
| technological progress is responsible for every creature
| comfort I have experienced in my entire life. People
| deserve to make livings off of those things even if they
| weren't solely responsible for the innovation.
| hokumguru wrote:
| I fully understand the hype but the initial scene with Sam
| feels _nothing_ like how any self respecting video producer
| would create. The jump cuts mid-sentence are extremely
| jarring, certainly not framed in any traditional sense, and
| he 's almost entirely out of focus.
|
| Points though for the completely expressionless line
| delivery, it completely nailed that.
| zendayawins6 wrote:
| It definitely doesnt look worse tbh. Its impressive stuff you
| cant get around that
| jayd16 wrote:
| I get what you're saying. I see it too. That said, a lot of
| people won't notice the flaws, especially with these fast,
| choppy cuts. By the time you realize the neck is way too long
| or whatever, its 2 cuts later.
| calmoo wrote:
| Anecdotally, I forgot multiple times that I was watching AI
| generated content, and my partner tuning in and out of it
| asked me a few times if we were watching the demo (as opposed
| to the real video). We are both pretty sensitive to slop too.
| I think something has flipped here.
| iLoveOncall wrote:
| I'm honestly convinced that half (if not a lot more) of the
| people commenting on HackerNews posts about OpenAI,
| Anthropic and similar companies are bots / employees paid
| by the company in question.
|
| You have to be unfathomably disingenuous to watch 2 minutes
| of video and "forget" that you are watching AI generated
| content, especially when it is glaringly obvious that it is
| such, and even more so when you claim to be "sensitive to
| slop".
|
| And funnily enough, looking at your comment history pretty
| much confirms this belief. And looking at another response
| I received (https://news.ycombinator.com/item?id=45431051),
| new account that just got created an hour ago solely to
| praise the result. Doesn't look suspicious at all.
| calmoo wrote:
| Let me tell you, I wish I was paid by OpenAI, but
| unfortunately I get paid a very average wage in a country
| where OpenAI do not operate.
|
| You can look at my account history and see that there is
| absolutely no indication that I am paid for, or a bot, I
| just happen to be relatively enthusiastic and bullish on
| AI capabilities. My account is 5 years old.
|
| I do not think this product from OpenAI will be good for
| the world, I think it is mostly a very bad thing. I just
| think it is impressive.
|
| I encourage you to try make the most generous
| interpretation of my comment and try to consider that
| other people can have sometimes wildly different
| experiences to you, however hard that is to believe. The
| alternative is to accuse people of being paid for
| commentors, which benefits no one and just isolates you
| further.
|
| To add more nuance to my original comment, yes if you
| look closely at the videos it is clear that they are AI
| generated, but as I was watching I 'lapsed' out of
| attention a few times and had to remind myself what I was
| looking at was generated. If I stopped and rewound a few
| times, yes of course I could notice the slop. But my
| point is that it's the first time I have lost the "I am
| watching slop" thought while watching AI video.
|
| Hope that clears things up for you.
|
| Edit: I just looked through your comment history, you are
| an extreme AI perma bear, and that's ok, that doesn't
| mean you're a paid for anti-ai lobbyist. Also every
| single comment you have made is extremely negative and
| inflammatory.
|
| I encourage you to revise your communication methods and
| try to practice some kindness.
| VagabundoP wrote:
| I hate this vacant technology tbh. Every video feels like
| distilled advert mindless slop.
|
| There's still something off about the movements, faces and eyes.
| Gollum features.
| horhay wrote:
| So far the true progress it has made is getting textures right
| close up. It still fudges how skin looks like the more it pans
| away from the characters.
| mrcino wrote:
| So, this is the AI Slop generator for the AI SlipSlop that Altman
| has announced lately.
|
| Brave new internet, where humans are not needed for any "social"
| media anymore, AI will generate slop for bots without any human
| interaction in an endless cycle.
| carabiner wrote:
| CEO of Loopt makes a cameo at 1:28 in the youtube vid.
| mclightning wrote:
| It is very underwhelming. It seems like a step backward. Scam
| altman should be replaced before he runs the company to
| bankruptcy.
| iLoveOncall wrote:
| Show me a coherent video that lasts more than 5 seconds and was
| generated with the model and maybe I'll start to care.
| the_duke wrote:
| I haven't seen comments regarding a big factor here:
|
| It seems like OpenAI is trying to turn Sora into a social network
| - TikTok but AI.
|
| The webapp is heavily geared towards consumption, with a feed as
| the entry point, liking and commenting for posts, and user
| profiles having a prominent role.
|
| The creation aspect seems about as important as on Instagram,
| TikTok etc - easily available, but not the primary focus.
|
| Generated videos are very short, with minimal controls. The only
| selectable option is picking between landscape and portrait mode.
|
| There is no mention or attempt to move towards long form videos,
| storylines, advanced editing/controls/etc, like others in this
| space (eg Google Flow).
|
| Seems like they want to turn this into AITok.
|
| Edit: regarding accurate physics ... check out these two videos
| below...
|
| To be fair, Veo fails miserably with those prompts also.
|
| https://sora.chatgpt.com/p/s_68dc32c7ddb081919e0f38d8e006163...
|
| https://sora.chatgpt.com/p/s_68dc3339c26881918e45f61d9312e95...
|
| Veo:
|
| https://veo-balldrop.wasmer.app/ballroll.mp4
|
| https://veo-balldrop.wasmer.app/balldrop.mp4
|
| Couldn't help but mock them a little, here is a bit of fun... the
| prompt adherence is pretty good, at least.
|
| NOTE: there are plenty of quite impressive videos being posted,
| and a lot of horrible ones also.
| ch4s3 wrote:
| That seems like an awful use of technology like this. I would
| imagine they mean to use that for serving ads, but how do you
| even generate conversations with ai slop plus product
| placements? I could see it working sometimes but I doubt it
| scales.
| micromacrofoot wrote:
| > slop plus product placements
|
| social media was heading this way before AI
| Computer0 wrote:
| Are users of the $20 tier really going to have to deal with
| that obnoxious bouncing watermark I wonder? The previous
| watermark could be cropped, but I often didn't feel the need to
| as I use it for fun, but that would make me not want to show
| anyone.
| bonoboTP wrote:
| Meta did the same recently:
| https://about.fb.com/news/2025/09/introducing-vibes-ai-video...
| echelon wrote:
| I posit this is the real story.
|
| OpenAI did not stealthily release Sora 2 to the image and video
| ELO ranking leaderboards ahead of time as is now somewhat
| tradition.
|
| This model is probably designed to run fast and cheap as a
| social play. Emphasis on putting you and your friends into
| popular franchises and IPs.
|
| OpenAI probably has a totally different model for their
| Hollywood-grade VFX. One that's too expensive to offer $20/mo
| consumers.
|
| - - - - -
|
| EDIT:
|
| Oh my god, OpenAI literally just disrupted TikTok:
|
| https://x.com/GabrielPeterss4/status/1973071380842229781
|
| https://x.com/GabrielPeterss4/status/1973122324984693113
|
| https://x.com/GabrielPeterss4/status/1973121891926942103
|
| https://x.com/GabrielPeterss4/status/1973120058907041902
| (potentially dangerous ... )
|
| https://x.com/GabrielPeterss4/status/1973111654524264763
|
| https://x.com/GabrielPeterss4/status/1973090475486879818
|
| https://x.com/GabrielPeterss4/status/1973110596825653720 (is
| this the same model? It doesn't look like it.)
|
| https://x.com/GabrielPeterss4/status/1973096194508251321
|
| https://x.com/GabrielPeterss4/status/1973086729281347650
|
| https://x.com/GabrielPeterss4/status/1973088038851932522 (this
| is truly something only kids will love)
|
| https://x.com/GabrielPeterss4/status/1973087595967201449
|
| https://x.com/GabrielPeterss4/status/1973077105903620504
|
| Holy shit!
|
| This is 100% the future of what kids will do. This is
| incredible for short form vertical video.
|
| It doesn't need to look good, it just needs to let you tell
| incredible stories with people and things you care about.
|
| This is way better than Meta's social video app.
| Gud wrote:
| Why would I want to watch any of this?
| jahsome wrote:
| You might not want to. It's definitely not appealing to me
| in any way shape or form.
|
| The younger generations however will likely gobble it right
| up. I try not to judge because folks said the same thing
| about Nintendo when I was young.
| kiririn7 wrote:
| i hate being ascended beings living above society
| BizarroLand wrote:
| STOP_HAVING_FUN.gif
| zain37 wrote:
| Just like how the hype on Ghibli art styles via ChatGPT
| died, same will probably happen here
| password54321 wrote:
| Touch grass. This is nothing but cringe that I wouldn't wish
| upon children.
| the_duke wrote:
| Kids are going to absolutely love this.
| smrtinsert wrote:
| I agree with the idea that they will like it, but I don't
| think it will look anything like this. I imagine native AI
| generations willl produce content will probably be
| instrutable to anyone older, requiring meme translations.
| Maybe a channel can be an AI decipherer. Hah.
|
| I've long thought that AI will force new distribution methods
| because old media is so markedly against it... Maybe this is
| another Netflix vs Blockbuster moment.
| motoxpro wrote:
| People in this thread saying that this is the kind of content
| kids like should go on tiktok for a sec. This is not at all
| what young people watch, it's just bad content, and
| misunderstanding that feels out of touch.
| micromacrofoot wrote:
| Yes, this is actually what they're trying to do. Internally
| they've been working on a social network for a while but it's
| kind of languishing.
| ares623 wrote:
| I'm sure all the influencers pushing AI art will be thrilled
| about this.
| rvz wrote:
| I bet xAI and X will likely relaunch Vine with AI videos as a
| competitor to Sora 2.
| crucialfelix wrote:
| > It seems like OpenAI is trying to turn Sora into a social
| network - TikTok but AI.
|
| That's a direct copy of what Midjourney has done already.
|
| https://www.midjourney.com/explore?tab=videos
|
| Many people are just playing with images and the distinctive
| styles that Midjourney (the model) seems to have developed.
| It's also trained by ratings and people's interactions.
|
| When you make images you can dial down the "aesthetic".
| zain37 wrote:
| This app might top the charts via hype initially but I can't
| see why someone would stick with it long-term compared to
| other alternatives. Plus creators would have to pay soon to
| make these videos, what are they getting back? Unless they
| can make money via this
| kelvinjps10 wrote:
| Free access to openAI tools?
| j45 wrote:
| That's one way to build a database of verified Gen AI content
| to help filter it out.
| some-guy wrote:
| My opinion is that unless there is some insane breakthrough in
| power efficiency with video generation, or if energy costs go
| down to zero, there is no way such a thing actually becomes
| profitable at the scale of scrolling TikTok. It is far more
| power efficient (and cheaper) to have people post their own
| content.
| a1371 wrote:
| When they launched Sora, one of the first things people did was
| rendering a person holding a cardboard with a message on it. It
| started by asking for features and eventually turned into
| people responding to each other.
|
| One conversation I remember was complaining about people who
| constantly want AI pictures of anime feet.
|
| I think OpenAI is just responding to the users.
| razodactyl wrote:
| How dare you be critical about a service offering in favour of
| a better end-user experience! /s
| ionwake wrote:
| I think HN is too political like this tech is clearly amazing and
| it's great they shipped it there should be more props even if
| it's a billion dollar company.
| Voloskaya wrote:
| Yes the tech is amazing. But tech is not everything, after 20
| years of social media, its pretty clear to everyone that those
| things can have large long term impact both positive and
| negative for society, discussing the potential impacts of the
| tech is not being "political", its just being interested in the
| future.
| unfitted2545 wrote:
| There's a great lyric from ELUCID I think about when people say
| stuff like this:
|
| > I don't have the privilege to think everything ain't
| political
| doikor wrote:
| Does this survive panning the camera away for 5 to 10 seconds and
| then back? Or basic conversation scene with the camera cutting
| between being located behind either speaker once every few
| seconds?
|
| Basically proper working persistence of the scene.
| bsenftner wrote:
| Dude, this generation of AI video models are just starting to
| have basic camera production terms understood, and then it is
| exactly like LLM generation: it's a pull of a slot machine arm;
| you might get what you want, but that's "winning" and the slot
| machine only gives out winners one in every 100 pulls. Every
| possible thing that could not be right happens.
|
| For example, I'm working with a walking and talking character
| at this time using multiple AI video models and systems.
| Generated clips any length longer than 8 seconds risk rapid
| quality loss, but _sometimes_ you can get up to 12-19 seconds
| without the generation breaking down. That means one needs to
| simulate a multiple camera shoot on a stage, so you can cut
| around the character(s) and create a longer sequence. But now
| you need to have multiple views of the same location to place
| your character(s) into - and current AI models can 't reliably
| give you a "different angled views" of an environment. We just
| got consistent different views of characters, and it'll be
| another period until environments can be generally examined
| from any view. BUT, that's if people realize this is not in the
| models yet, and so far people are so fascinated by the fantasy
| violence and sexual content they can make nobody realizes you
| cannot simply "look left and right" in any of these models and
| that even works with consistency or reliability. There are
| workarounds, like creating one's entire set and environments in
| 3D models, for use as the backgrounds and starting frames, but
| that's now 3D media production + AI, and none of the AI tools
| generate media that even has alpha channels, and a lot of
| similar incompatibilities like that.
| carrozo wrote:
| Sora 2: Sloppy Seconds
| CSMastermind wrote:
| Anyone have an invite they want to share with me lol.
| apetresc wrote:
| If anyone is feeling generous with one of their four invite
| codes, I'd really appreciate it. I'm at adrian@apetre.sc.
| TheAceOfHearts wrote:
| Really impressive engineering work. The videos have gotten good
| enough that they can grab your attention and trigger a strong
| uncanny valley feeling.
|
| I think OpenAI is actually doing a great job at easing people
| into these new technologies. It's not such a huge leap in
| capabilities that it's shocking, and it helps people acclimate
| for what's coming. This version is still limited but you can tell
| that in another generation or two it's going to break through
| some major capabilities threshold.
|
| To give a comparison: in the LLM model space, the big
| capabilities threshold event for me came with the release of
| Gemini 2.5 Pro. The models before that were good in various ways,
| but that was the first model that felt truly magical.
|
| From a creative perspective, it would be ideal if you could first
| generate a fixed set of assets, locations, and objects, which are
| then combined and used to bring multiple scenes to life while
| providing stronger continuity guarantees.
| NoahZuniga wrote:
| TTS is horrible compared to Google's veo 3
| neilv wrote:
| > _And we 're introducing Cameo, giving you the power to step
| into any world or scene, and letting your friends cast you in
| theirs._
|
| How much are they (and providers of similar tools) going to be
| able to keep _anyone_ from putting _anyone else_ in a video,
| shown doing and saying whatever the tool user wants?
|
| Will some only protect politicians and celebrities? Will the
| less-famous/less-powerful of us be harassed, defamed, exploited,
| scammed, etc.?
| notatoad wrote:
| it seems like this is basically youtube's ContentID, but for
| your face. as long as you upload your "cameo" aka facial scan
| to them, they can recognize and control the generation of
| videos with it. if you don't give them your face, then they
| can't/won't.
|
| "Consent-based likeness. Our goal is to place you in control of
| your likeness end-to-end with Sora. We have guardrails intended
| to ensure that your audio and image likeness are used with your
| consent, via cameos. Only you decide who can use your cameo,
| and you can revoke access at any time. We also take measures to
| block depictions of public figures (except those using the
| cameos feature, of course). Videos that include your cameo--
| including drafts created by other users--are always visible to
| you. This lets you easily review and delete (and, if needed,
| report) any videos featuring your cameo. We also apply extra
| safety guardrails to any video with a cameo, and you can even
| set preferences for how your cameo behaves--for example,
| requesting that it always wears a fedora."
| neilv wrote:
| If this company's guardrails end up sufficiently working well
| in practice (note phrases like "intended", "take measures",
| and "preferences...requested", on things they can't do
| 100%)... there will be weak links elsewhere, letting similar
| computation be performed without sufficiently effective
| guardrails against abuse?
|
| How do we prepare for this? Societal adjustment only (e.g.,
| disbelieving defamatory video, accepting what pervs will do)?
| Establishing a common base of cultural expectations for
| conduct? Increasing deterrence for abusers?
| felixakiragreen wrote:
| Brilliant.
|
| Until you have 2 people that are near identical. They don't
| even have to be twins, there are plenty of examples where
| people can't even tell other people apart. How is an AI going
| to do it?
|
| You don't own your likeness. It's not intellectual property.
| It's a constantly changing representation of a biological
| being. It can't even be absolutely defined-- it's always
| subject to the way in which it was captured. Does a person
| own their likeness for all time? Or only their current
| likeness? What about more abstract representations of their
| likeness?
|
| The can of worms OpenAI is opening by going down this path is
| wild. We're not current able to solve such a complex issue.
| We can't even distinguish robots from humans on the internet.
| mvdtnz wrote:
| I'm an identical twin so immediately I can see a pretty
| stupid obvious problem with this.
| colesantiago wrote:
| Basically deepfakes for everyone.
| echelon wrote:
| Honestly this is the safest possible outcome.
|
| If Deepfakes remain the tools of nation state actors,
| laypeople will be easily fooled.
|
| If Deepfakes are available on your iPhone and within TikTok,
| everyone will just ask "Is it Photoshop?" for every shred of
| doubt. (In fact, I already see people saying, "This looks
| like AI".)
|
| This is good. Normalize the magic until it isn't magic
| anymore.
|
| People will get it. They're smart. They just need exposure.
| colesantiago wrote:
| > People will get it. They're smart. They just need
| exposure.
|
| I really doubt this.
|
| If you are in the creative field, your work will just be
| reduced to "is this slop?" or "fixed it!" with a low effort
| AI generated work of your original work (fuck copyright
| right?).
|
| I already see artists battling and fighting putting out
| their best non AI work only for their audience to question
| if it is real and they lose the impressiveness.
|
| This just already undermines creators who don't use AI
| generated stuff.
|
| But who cares about them right? "it is the future" and it
| is most _definitely_ AGI for them.
|
| But then again, the starving artist never really made any
| money and this ensures that the artform stays dead.
| pton_xd wrote:
| > People will get it. They're smart. They just need
| exposure.
|
| It's either this, or the opposite (eg, misinformation needs
| to be censored). Seems like we as a society can't quite
| make up our mind on which approach to take.
| rhetocj23 wrote:
| Ah the great trade off that comes with little to no
| regulation.
| rvz wrote:
| 12,000+ "AI startups" have been obliterated.
| bgwalter wrote:
| What is the target market for this? The videos are not good
| enough for YouTube. They are unrealistic, nauseating and dorky.
| Already now any YouTube video that contains a hint of "AI"
| attracts hundreds of scathing comments. People do not want this.
|
| Let me guess, the ultimate market will be teenagers "creating" a
| Skibidi Toilet and cheap TikTok propaganda videos which promote
| Gazan ocean front properties.
| LarsDu88 wrote:
| I really hope they have more granular APIs around this.
|
| One use case I'm really excited about is simply making animated
| sprites and rotational transformations of artwork using these
| videogen models, but unlike with local open models, they never
| seem to expose things like depth estimation output heads, aspect
| ratio alteration, or other things that would actually make these
| useful tools beyond shortform content generation.
| jp57 wrote:
| Prediction: we'll see at least one Sora-generated commercial at
| the Super Bowl this year.
| vahid4m wrote:
| While the quality of what I'm seeing is very nice for AI
| generated content (I still can't believe it) but the fact thay
| they are mostly showing short clips and not a long connected
| consistent video makes it less impressive.
| squidsoup wrote:
| A little tangential to this announcement, but is anyone aware of
| any clean/ethical models for AI video or image generation (i.e.
| not trained on copyright work?) that are available publicly?
| egeres wrote:
| I wonder how this will affect the large cinema production
| companies (Disney, WB, Universal, Sony, Paramount, 20th
| century...). The global film market share was estimated to be
| 100B in 2023. If the production cost of high FX movies like
| Avengers Infinity War goes down from 300M$ to just 10K$ in a
| couple of years, will companies like Disney restrain themselves
| to just release a few epic movies per year? Or will we be flooded
| with tons of slop? If this kind of AI content keeps getting
| better, how will movies sustain our attention and feel 'special'?
| Will people not care if an actor is AI or real?
| ashu1461 wrote:
| This is a good comparison thread of capabilities of sora vs sora
| 2
|
| https://x.com/mattshumer_/status/1973085321928515783
| seydor wrote:
| Since Agi is cancelled, at least we have shopping and endless
| video
| clgeoio wrote:
| > Concerns about doomscrolling, addiction, isolation, and RL-
| sloptimized feeds are top of mind--here is what we are doing
| about it.
|
| > We are giving users the tools and optionality to be in control
| of what they see on the feed. Using OpenAI's existing large
| language models, we have developed a new class of recommender
| algorithms that can be instructed through natural language. We
| also have built-in mechanisms to periodically poll users on their
| wellbeing and proactively give them the option to adjust their
| feed.
|
| So, nothing? I can see this being generated and then reposted to
| TikTok, Meta, etc for likes and engagement.
| alberth wrote:
| Why do you have to download an app to use Sora 2 (vs it being
| available on the web like ChatGPT)?
| samuelfekete wrote:
| This is a step towards a constant stream of hyper-personalised AI
| generated content optimised for max dopamine.
| taberiand wrote:
| The Torment Nexus is a Skinner box
| fersarr wrote:
| Only iphone...
| nycdatasci wrote:
| What makes TikTok fun is seeing actual people do crazy stuff.
| Sora 2 could synthesize someone hitting five full-court shots in
| a row, but it wouldn't be inspiring or engaging. How will this be
| different than music-generating AI like Suno, which doesn't have
| widespread adoption despite incredible capabilities?
| heldrida wrote:
| It's hard to believe, but some people enjoy. On the other hand,
| some popular content on TikTok is probably worse than AI
| generated content and that's another problem...
| dolebirchwood wrote:
| This makes me less excited about the future of video, not more.
|
| It's technically impressive, but all so very soulless.
|
| When everything fake feels real, will everything real feel fake?
| nalimtasseb wrote:
| Truly wonder if there will be some kind of renaissance in the
| video making domain when all settles down and this becomes the
| new normal.
| rhetocj23 wrote:
| Tastes and preferences are dynamic. It will certainly happen.
| bsenftner wrote:
| The ease of creating visually titillating media, coupled with
| the difficultly of consistency works against the creation of
| narrative media. I sure hope we don't get a generation of
| non-narrative beautiful slop.
| amelius wrote:
| Nicely cherry-picked.
| ezomode wrote:
| full-on productisation effort -> no AGI in sight
| Josh5 wrote:
| Everyone has the widest eyes in these Sora videos.
| FullMetul wrote:
| Maybe by Sora 3 they will have scene consistency. Gah it's so
| jarring to me that the poll the racing ducks are in just randomly
| changes. My brain can tell it's not consistent scene to scene and
| feels so jank.
| groos wrote:
| What is the point? Who wants to watch these videos?
| Havoc wrote:
| That sure seems to be getting close to something usable for
| movies...kinda.
|
| Sam looks weirdly like Cillian Murphy in Oppenheimer in some
| shots. I wonder whether there was dataset bleedover from that.
| yahoozoo wrote:
| Sam still pretending they're close to AGI in the trailer lmao
| cogman10 wrote:
| I've seen a lot of "this is impressive" but I'm not really seeing
| it. This looks to suffer from all the same continuity problems
| other AI videos suffer from.
|
| What am I looking at that's super technically impressive here?
| The clips look nice, but from one cut to the next there's a lot
| of obvious differences (usually in the background, sometimes in
| the foreground).
| paulcole wrote:
| As a gauge for how seriously I should take your critique:
|
| How many hours a week are you actively using AI tools yourself?
|
| What percentage of public comments that you've made about AI
| tools have been skeptical or critical?
| cogman10 wrote:
| > How many hours a week are you actively using AI tools
| yourself?
|
| 2 or 3. Mostly LLMs to check code.
|
| > What percentage of public comments that you've made about
| AI tools have been skeptical or critical?
|
| Probably around 90%.
|
| So sell me. Why is this super impressive? I'm happy to admit
| that I'm pretty pessimistic about AI.
|
| I have an eye for continuity issues, they are pretty obvious
| to me. Am I just too focused on that sort of a thing?
| paulcole wrote:
| > Why is this super impressive?
|
| It's fucking video made by a computer after you type a
| sentence. I don't get how this isn't insanely super
| impressive.
| umrashrf wrote:
| hey @simoncion looks like they are doing this for self-promotion
| that's against the site's guidelines
| dcreater wrote:
| Matrix here we come!
| Aeolun wrote:
| Clicking a link on the OpenAI dashboard and beeing greeted with a
| full page of scandily clad women was certainly not what I
| expected to see when opening Sora..
___________________________________________________________________
(page generated 2025-09-30 23:00 UTC)