Live data from Hacker News

Sora: Creating video from text

openai.com

201–210 of 1001 posts

Re: Sora: Creating video from text

#201

OpenAI demonstrating the size of their moat. How many multi-million-dollar funded startups did this just absolutely obsolete? This is so, so, so much better than every other generative video AI we've seen. Most of those were basically a still image with a very slowly moving background. This is not that. Sam is probably going to get his $7T if he keeps this up, and when he does everybody else will be locked out foreve…

Even with 7 trillion, he is still going to need a national grid that can supply the power for the compute.

There is a lot that has to planned and put in place now to get there.

As for people that have opted out of life. We would have a better world if we started encouraging more dreamers/doers like out of the movie Tomorrowland.

Re: Sora: Creating video from text

#202

OpenAI demonstrating the size of their moat. How many multi-million-dollar funded startups did this just absolutely obsolete? This is so, so, so much better than every other generative video AI we've seen. Most of those were basically a still image with a very slowly moving background. This is not that. Sam is probably going to get his $7T if he keeps this up, and when he does everybody else will be locked out foreve…

> OpenAI demonstrating the size of their moat. How many multi-million-dollar funded startups did this just absolutely obsolete? For posterity since the term has been misused lately, having a very good product isn't a moat in the business sense. There's nothing stopping a competitor from creating a similar product (even if it's difficult), and there's nothing currently stopping OpenAI's users from switching from using…

Being (a) first and (b) good enough is a moat. Nothing stopped people from switching from google to bing all these years other than not having any reason to.

Re: Sora: Creating video from text

#204

This is leaps and bounds beyond anything out there, including both public models like SVD 1.1 and Pika Labs' / Runway's models. Incredible.

Let's hold our breath. Those are specifically crafted hand-picked good videos, where there wasn't any requirement but "write a generic prompt and pick something that looks good", with no particular requirements. Which is very different from the actual process where you have a very specific idea and want the machine to make it happen. DALL-E presentation also looked cool and everyone was stoked about it. Now that we k…

> Stable Diffusion is still the go-to solution. I strongly suspect the same thing with Sora.

Sure, for people who want detailed control with AI-generated video, workflows built around SD + AnimateDiff, Stable Video Diffusion, MotionDiff, etc., are still going to beat Sora for the immediate future, and OpenAI's approach structurally isn't as friendly to developing a broad ecosystem adding power on top of the base models.

OTOH, the basic simple prompt-to-video capacity of Sora now is good enough for some uses, and where detailed control is not essential that space is going to keep expanding -- one question is how much their plans for safety checking (which they state will apply both to the prompt and every frame of output) will cripple this versus alternatives, and how much the regulatory environment will or won't make it possible to compete with that.

Re: Sora: Creating video from text

#205
post #87

This is going to make the latest election really interesting (and scary). Is anyone working to ensure a faked video of Biden that looks plausible but is AI generated doesn't get significant traction at a critical moment of the election?

That just doesn't seem like a plausible scenario to me. Obviously, if such thing happened, Biden would have an alibi, since it's known where he is at all times.

The people who already hate Biden, probably already think he's doing some weird shady stuff, and would point to some conspiracy. The people who like Biden, would accept the alibi.

Ultimately it wouldn't move the needle.

What is concerning, is the technology being used against a regular person, who may not have an alibi.

Re: Sora: Creating video from text

#206
Obviously incredibly cool, but it seems that people are incredibly overstating the applications of this.

Realistically, how do you fit this into a movie, a TV show, or a game? You write a text prompt, get a scene, and then everything is gone—the characters, props, rooms, buildings, environments, etc. won’t carry over to the next prompt.

Re: Sora: Creating video from text

#208

This is insane. Even though there are open-source models, I think this is too dangerous to release to the public. If someone would've uploaded that Tokyo video to youtube, and told me it was a drone.. I would've believed them. All "proof" we have can be contested or fabricated.

This is what lots of folks said about image generation. Which is now in many ways “solved”. And society has easily adapted to it. The same will happen with video generation. The reality is that people are a lot more resourceful / smarter than a lot of us think. And the ones who aren’t have been fooled long before this tech came around.

In what ways has image generation been solved? Prompt blocking is about the only real effort I can think of, which will mean nothing once open source models reach the same fidelity.

Re: Sora: Creating video from text

#210
post #162

OpenAI demonstrating the size of their moat. How many multi-million-dollar funded startups did this just absolutely obsolete? This is so, so, so much better than every other generative video AI we've seen. Most of those were basically a still image with a very slowly moving background. This is not that. Sam is probably going to get his $7T if he keeps this up, and when he does everybody else will be locked out foreve…

OpenAI's moat is (a) talent (b) access to compute (c) no fear of using whatever data they can get. On the other hand, I think these moats will be destroyed as soon as anyone finds a drastically more efficient (compute- and data-wise) way to train LLMs. Biology would suggest that it doesn't take $100 million worth of GPUs and exaflops of compute to achieve the intelligence of a human. (Of course it is possible that at…

Biology literally took a planet sized genetic algorithm with nanomachines a couple Billion years to get to this point.
Post reply on HN