Live data from Hacker News

Sora: Creating video from text

openai.com

351–360 of 1001 posts

Re: Sora: Creating video from text

#351

This is leaps and bounds beyond anything out there, including both public models like SVD 1.1 and Pika Labs' / Runway's models. Incredible.

Let's hold our breath. Those are specifically crafted hand-picked good videos, where there wasn't any requirement but "write a generic prompt and pick something that looks good", with no particular requirements. Which is very different from the actual process where you have a very specific idea and want the machine to make it happen. DALL-E presentation also looked cool and everyone was stoked about it. Now that we k…

Would love to see handpicked videos from competitors that can hold their own against what SORA is capable of

Re: Sora: Creating video from text

#352

OpenAI demonstrating the size of their moat. How many multi-million-dollar funded startups did this just absolutely obsolete? This is so, so, so much better than every other generative video AI we've seen. Most of those were basically a still image with a very slowly moving background. This is not that. Sam is probably going to get his $7T if he keeps this up, and when he does everybody else will be locked out foreve…

We say 7T$ as if it’s nothing, am I the only one shocked by the sum we are talking about? This is close to what BlackRock is managing!

I'm fairly sure $7T is a speculation bubble, and that's going to pop like all bubbles pop. It's the combined GDP of Japan and Canada. It's too big for an investment.

It's not necessarily too big for a valuation, as a sufficiently capable AI is an economic power in its own right: I previously guessed, and even despite its flaws would continue to guess within the domain of software development at least, that the initial ChatGPT model was about as economically valuable to each user as an industrial placement student, and when I was one of those I was earning about £1.7k/month when adjusted for inflation, US$2.1k at current nominal exchange rates. 100 million users at that rate is $2.52e+12/year in economic productivity, and that's with the current chip supply and (my estimate of) the productivity of a year-old model — and everyone knows that this sector is limited by the chips, and that $7T investment story is supposed to be about improving the supply of those chips.

Re: Sora: Creating video from text

#353
post #24

Yeah, you just can't let all media, all the cost and hard work of millions of photographers, animators, filmmakers, etc be completely consumed and devalued by one company just because it's a very cool technical trick. The more powerful these services become the more obvious that will be. What OpenAI does is amazing, but they obviously cannot be allowed to capture the value of every piece of media ever created — it'll…

Do you feel the same about the hard work of knocker-uppers having been devalued by the invention of the alarm clock? Or is it just the (relatively) highly paid intellectual workers that "cannot be allowed" to be replaced with machines?

Re: Sora: Creating video from text

#355

I'm not sure about others, but I'm extremely unnerved about how OpenAI just throws these innovations out with zero foreshadowing - it's crazy how the world's potentially most life-changing company operates with the secrecy of a black military program. I really wonder what's going to come out of the company and on what timeline.

That's what's mindblowing to me It doesn't feel like a slow incremental progress, the last AI videos I've seen were terrible Its like suddenly a huge jump in quality

It is a sudden jump in quality. A mere _month_ ago, this is what googles SOTA was: https://lumiere-video.github.io/

Re: Sora: Creating video from text

#356

Many might miss the key paragraph at the end: "Sora serves as a foundation for models that can understand and simulate the real world, a capability we believe will be an important milestone for achieving AGI." This also helps explain why the model is so good since it is trained to simulate the real world, as opposed to imitate the pixels. More importantly, its capabilities suggest AGI and general robotics could be cl…

What is latent space if not a representation of the real world?

Pretty sure many latent spaces are not trained to represent 3D motions and some detailed physics of the real world. Those in pure text LLMs, for example.

Re: Sora: Creating video from text

#358

In less than a few hours Gemini 1.5 is old news. Sam is doing live demos on Twitter while Google just released a blog. Didn't think Google would be the first of the Facebook, Apple, Google and Microsoft to get disrupted.

I mean, why would this make google look bad?

Gemini is catching up, so OpenAI needs a new venue to market itself to the investors. It is doing a soft pivoting if you ask me, now GPT4 is like not that special anymore.

Re: Sora: Creating video from text

#359

Many might miss the key paragraph at the end: "Sora serves as a foundation for models that can understand and simulate the real world, a capability we believe will be an important milestone for achieving AGI." This also helps explain why the model is so good since it is trained to simulate the real world, as opposed to imitate the pixels. More importantly, its capabilities suggest AGI and general robotics could be cl…

> "understand... the real world"

doing a lot of heavy lifting in this statement

Post reply on HN