Live data from Hacker News

Sora: Creating video from text

openai.com

71–80 of 1001 posts

Re: Sora: Creating video from text

#71
post #42

This is leaps and bounds beyond anything out there, including both public models like SVD 1.1 and Pika Labs' / Runway's models. Incredible.

Agreed. It's amazing how much of a head start OpenAI appears to have over everyone else. Even Microsoft who has access to everything OpenAI is doing. Only Microsoft could be given the keys to the kingdom and still not figure out how to open any doors with them.

[flagged]

Re: Sora: Creating video from text

#73
post #41

Countdown to when studios licensing this for "unlimited" episodes of your favorite series. There was Seinfeld "Nothing, Forever" AI parody, but once the models improve enough and are cheap enough to deploy, studios will license their content for real and just have endless seasons. Or even custom episodes. Imagine if every episode of a TV show was unique to the viewer.

I imagine it's not long before we see hyper-targeted commercials where the actors look like us, live in our city, etc.

Wow, I would imagine this being very effective in election campaigns (for better or for worse, probably for worse).

Re: Sora: Creating video from text

#74
post #42

This is leaps and bounds beyond anything out there, including both public models like SVD 1.1 and Pika Labs' / Runway's models. Incredible.

Agreed. It's amazing how much of a head start OpenAI appears to have over everyone else. Even Microsoft who has access to everything OpenAI is doing. Only Microsoft could be given the keys to the kingdom and still not figure out how to open any doors with them.

Many people say the same about Google/DeepMind.

Re: Sora: Creating video from text

#75
post #24

Yeah, you just can't let all media, all the cost and hard work of millions of photographers, animators, filmmakers, etc be completely consumed and devalued by one company just because it's a very cool technical trick. The more powerful these services become the more obvious that will be. What OpenAI does is amazing, but they obviously cannot be allowed to capture the value of every piece of media ever created — it'll…

>Yeah, you just can't let all media, all the cost and hard work of millions of photographers, animators, filmmakers, etc be completely consumed and devalued by one company just because it's a very cool technical trick.

Oh man, how I miss it when ice was hauled from the Arctic in boats.

Re: Sora: Creating video from text

#76
post #40

Visual sharpness at the expense of wider-scale coherence (see: sliding/floating walking woman in Tokyo demo or tiny people next to giant people in Lagos demo) seems to be a local optimum consistently achieved by today's SOTA models in all domains. This is neat and all but mostly just a toy. Everything I've seen has me convinced either we are optimizing the wrong loss functions or the architectures we have today are f…

>Visual sharpness at the expense of wider-scale coherence (see: sliding/floating walking woman in Tokyo demo or tiny people next to giant people in Lagos demo)

Wider-Scale coherence is still much better than previous models and has consistently been improving. It's not "visual sharpness at the expense of coherence". At worst, the models are learning wider-scale coherence slower.

Not everything is equally difficult to learn so it follows that some aspects will lag behind others. If coherence weren't improving you might have a point but it is so...

Re: Sora: Creating video from text

#77
post #24

Yeah, you just can't let all media, all the cost and hard work of millions of photographers, animators, filmmakers, etc be completely consumed and devalued by one company just because it's a very cool technical trick. The more powerful these services become the more obvious that will be. What OpenAI does is amazing, but they obviously cannot be allowed to capture the value of every piece of media ever created — it'll…

Sorry no. If there was even the remotest possibility that everyone could be brought to the table, none of these would even exist.

Training a massive model like this is a risk, and no one is going to take that risk without some reward. You can complain OpenAI is going to too much of the value, but its value that would have otherwise never existed. It's value.

Re: Sora: Creating video from text

#79

This is leaps and bounds beyond anything out there, including both public models like SVD 1.1 and Pika Labs' / Runway's models. Incredible.

Where is the training material for this coming from? The only resource I can think of that's broad enough for a general purpose video model is YouTube, but I can't imagine Google would allow a third party to scrape all of YT without putting up a fight.
Post reply on HN