Live data from Hacker News

Sora: Creating video from text

openai.com

821–830 of 1001 posts

Re: Sora: Creating video from text

#821

This is insane. But I'm impressed most of all by the quality of motion . I've quite simply never seen convincing computer-generated motion before . Just look at the way the wooly mammoths connect with the ground, and their lumbering mass feels real. Motion-capture works fine because that's real motion, but every time people try to animate humans and animals, even in big-budget CGI movies, it's always ultimately obvio…

Regarding CGI, I think it has became so good that you don’t know it’s CGI. Look at the dog in Guardians of the Galaxy 3. There’s a whole series on YouTube called “no cgi is really just invisible cgi” that I recommend watching.

And as with cgi, models like SORA will get better until you can’t tell reality apart. It's not there Yet, but an immense astonishingly breakthrough.

Re: Sora: Creating video from text

#822
post #406

This is insane. But I'm impressed most of all by the quality of motion . I've quite simply never seen convincing computer-generated motion before . Just look at the way the wooly mammoths connect with the ground, and their lumbering mass feels real. Motion-capture works fine because that's real motion, but every time people try to animate humans and animals, even in big-budget CGI movies, it's always ultimately obvio…

I disagree, just look at the legs of the woman in the first video. First she seems to be limping, than the legs rotate. The mammoth are totally uncanny for me as its both running and walking at the same time. Don't get me wrong, it is impressive. But I think many people will be very uncomfortable with such motion very quickly. Same story as the fingers before.

And further down the page the:

"The camera follows behind a white vintage SUV with a black roof": The letters clearly wobble inconsistently.

"A drone camera circles around a beautiful historic church built on a rocky outcropping along the Amalfi Coast": The woman in the white dress in the bottom left suddenly splits into multiple people like she was a single cell microbe multiplying.

Re: Sora: Creating video from text

#823

The Hollywood Reporter says many in the industry are very scared.[1] “I’ve heard a lot of people say they’re leaving film,” he says. “I’ve been thinking of where I can pivot to if I can’t make a living out of this anymore.” - a concept artist responsible for the look of the Hunger Games and some other films. "A study surveying 300 leaders across Hollywood, issued in January, reported that three-fourths of respondents…

The idea that this destroys the industry is overblown, because the film industry has already been dying since 2000's.

Hollywood is already destroyed. It is not the powerful entity it once was.

In terms of attention and time of entertainment, Youtube has already surpassed them.

This will create a multitude more YouTube creators that do not care about getting this right or making a living out of it. It will just take our attention all the same, away from the traditional Hollywood.

Yes there will still be great films and franchises, the industry is shrinking.

This is similar with Journalism saying that AI will destroy it. Well there was nothing to destroy because the a bunch of traditional newspapers already closed shop even before AI came.

Re: Sora: Creating video from text

#825

Watching these made me think, I'm going to want to go to the theatre a lot more in the future and see fellow humans in plays, lectures and concerts. Such achievements in technology must lead to cultural change. Look at how popular vinyl has become, why not theatre again.

The vinyl narrative is so whack.

https://www.riaa.com/u-s-sales-database/

At its peak, Inflation adjusted Vinyl Sales was $1.4billion in 1979. Then forward to the lowest sales in 2009 at $3.4million. So Vinyl has been so popular it grew to $8.5m by 2021.

That is just nostalgia, not cultural change pushed by the dystopia of AI.

Re: Sora: Creating video from text

#826

Has anyone else noticed the leg swap in Tokyo video at 0:14. I guess we are past uncanny, but I do wonder if these small artifacts will always be present in generated content. Also begs the question, if more and more children are introduced to media from young age and they are fed more and more with generated content, will they be able to feel "uncanniness" or become completely blunt to it. There's definitely interes…

Tangent to feeling numb to it - will it hinder children developing the understanding of physics, object permanence, etc. that our brains have?

There have been children, that reacted iritated, when they cannot swipe away real life objects. The idea is, to give kids enough real world experiences, so this does not happen.

Re: Sora: Creating video from text

#827

I think the implications go much further than just the image/video considerations. This model shows a very good (albeit not perfect) understanding of the physics of objects and relationships between them. The announcement mentions this several times. The OpenAI blog post lists "Archeologists discover a generic plastic chair in the desert, excavating and dusting it with great care." as one of the "failed" cases. But t…

Wouldn't having a good understanding of physics mean you know that a women doesn't slide down the road when she walks? Wouldn't it know that a woolly mammoth doesn't emit profuse amounts steam when walking on frozen snow? Wouldn't the model know that legs are solid objects in which other object cannot pass through? Maybe I'm missing the big picture here, but the above and all the weird spatial errors, like miniaturiz…

They could test this by trying to generate the same image but set in New York, etc. I bet it would still be asain.

Re: Sora: Creating video from text

#828
In the future, we're not going to have common tv shows or movies. We'll have a constantly evolving stream of entertainment that's perfectly customized to the viewer's preferences in real time. This is just the first step.

Re: Sora: Creating video from text

#829
post #49

https://openai.com/sora?video=big-sur In this video, there's extremely consistent geometry as the camera moves, but the texture of the trees/shrubs on the top of the cliff on the left seems to remain very flat, reminiscent of low-poly geometry in games. I wonder if this is an artifact of the way videos are generated. Is the model separating scene geometry from camera? Maybe some sort of video-NeRF or Gaussian Splatti…

My vote is yes - some sort of intermediate representation is involved. It just seems unbelievable that it's end-to-end with 2D frames...

Re: Sora: Creating video from text

#830

Watching these made me think, I'm going to want to go to the theatre a lot more in the future and see fellow humans in plays, lectures and concerts. Such achievements in technology must lead to cultural change. Look at how popular vinyl has become, why not theatre again.

The vinyl narrative is so whack. https://www.riaa.com/u-s-sales-database/ At its peak, Inflation adjusted Vinyl Sales was $1.4billion in 1979. Then forward to the lowest sales in 2009 at $3.4million. So Vinyl has been so popular it grew to $8.5m by 2021. That is just nostalgia, not cultural change pushed by the dystopia of AI.

Why is my 14 year old niece now collecting vinyl? I can guarantee it's not nostalgia. There's obviously more at play there even when acknowledging your point about relative market size.
Post reply on HN