Live data from Hacker News

Sora: Creating video from text

openai.com

471–480 of 1001 posts

Re: Sora: Creating video from text

#471

Does anyone know how to handle the depression/doom one feels with these updates? Yes, it's a great technical achievement, but I just worry for the future. We don't have good social safety nets, and we aren't close to UBI. It's difficult for me to see that happen unless something drastic changes. I'm also afraid of one company just having so much power. How does anyone compete?

UBI was tested during 2020 on a nearly global scale. In the US, the CARES Act which provided stimulus checks for every tax-paying US citizen as well as extensions to unemployment was essentially a giant UBI experiment. Not for AI, but for a giant shift in economic activity where many individuals became unemployed nonetheless.

https://en.wikipedia.org/wiki/CARES_Act

EDIT: For the downvoters, yes components of CARES was in fact inspired by UBI:

https://www.cnbc.com/2020/03/13/andrew-yang-aoc-free-ubi-cas...

https://www.businessinsider.com/coronavirus-aoc-demands-univ...

Re: Sora: Creating video from text

#472
post #111
post #49

https://openai.com/sora?video=big-sur In this video, there's extremely consistent geometry as the camera moves, but the texture of the trees/shrubs on the top of the cliff on the left seems to remain very flat, reminiscent of low-poly geometry in games. I wonder if this is an artifact of the way videos are generated. Is the model separating scene geometry from camera? Maybe some sort of video-NeRF or Gaussian Splatti…

Curious about what current SotA is on physics-infusing generation. Anyone have paper links? OpenAi has a few details: >> The current model has weaknesses. It may struggle with accurately simulating the physics of a complex scene, and may not understand specific instances of cause and effect. For example, a person might take a bite out of a cookie, but afterward, the cookie may not have a bite mark. >> Similar to GPT…

On the announcement page, it specifically says Sora does not understand physics

Re: Sora: Creating video from text

#473

This is insane. But I'm impressed most of all by the quality of motion . I've quite simply never seen convincing computer-generated motion before . Just look at the way the wooly mammoths connect with the ground, and their lumbering mass feels real. Motion-capture works fine because that's real motion, but every time people try to animate humans and animals, even in big-budget CGI movies, it's always ultimately obvio…

I'm not sure I feel the same way about the mammoths - and the billowing snow makes no sense as someone who grew up in a snowy area. If the snow was powder maybe but that's not what's depicted on the ground.

Re: Sora: Creating video from text

#474

Obviously incredibly cool, but it seems that people are incredibly overstating the applications of this. Realistically, how do you fit this into a movie, a TV show, or a game? You write a text prompt, get a scene, and then everything is gone—the characters, props, rooms, buildings, environments, etc. won’t carry over to the next prompt.

It doesn't need to replace the whole movie You could use it for stuff like wide shots, close ups, random CG shots, rapid cut shots, stuff where you just cut to it once and don't need multiple angles To me it seem most useful for advertising where a lot of times they only show something once, like a montage

And it would be magic for storyboarding. This would be such a useful tool for a director to iterate on a shot and then communicate that to the team

Re: Sora: Creating video from text

#475

Obviously incredibly cool, but it seems that people are incredibly overstating the applications of this. Realistically, how do you fit this into a movie, a TV show, or a game? You write a text prompt, get a scene, and then everything is gone—the characters, props, rooms, buildings, environments, etc. won’t carry over to the next prompt.

[deleted]

Re: Sora: Creating video from text

#476

Countdown to when studios licensing this for "unlimited" episodes of your favorite series. There was Seinfeld "Nothing, Forever" AI parody, but once the models improve enough and are cheap enough to deploy, studios will license their content for real and just have endless seasons. Or even custom episodes. Imagine if every episode of a TV show was unique to the viewer.

One understated aspect of AI Seinfeld is that it took many steps to differentiate it from the actual Seinfeld and create its own identity, such as the 144p visual filter and the random microwave. Those tweaks added to its charm. If someone tried to do AI Seinfeld again in 2024, many would criticze it for not being realistic enough now that the tools to do so are now available.

I assume you would still be able to do that, just better? Like pixel art. Super Mario Bros. 3 look great despite being 36 years old. Contrast this with 3D games for the original PlayStation that have aged poorly.

Re: Sora: Creating video from text

#477
post #49

https://openai.com/sora?video=big-sur In this video, there's extremely consistent geometry as the camera moves, but the texture of the trees/shrubs on the top of the cliff on the left seems to remain very flat, reminiscent of low-poly geometry in games. I wonder if this is an artifact of the way videos are generated. Is the model separating scene geometry from camera? Maybe some sort of video-NeRF or Gaussian Splatti…

Wow, yeah I didn't notice it at first, but looking at the rocks in the background is actually nauseating

Re: Sora: Creating video from text

#479

Did anyone else feel motion sickness or nausea watching some of these videos? In some of the videos with some panning or rotating motion, i felt some nausea like sickness effect. I guess its because some details were changing while in motion and I was unable to keep track or focus anything in particular. Effect was stronger in some videos.

Perfect fit for VR.
Post reply on HN