Live data from Hacker News

Sora: Creating video from text

openai.com

941–950 of 1001 posts

Re: Sora: Creating video from text

#941
post #873

This is both amazing and saddening to me. All our cultural legacy is being fed into a monstrous machine that gives no attribution to the original content with which it was fed, and so the creative industry seems to be in great danger. Creativity being automated while humans are forced to perform menial tasks for minimum wage doesn't seem like a great future and the geriatric political class has absolutely no clue how…

A part of the book Look to Windward by Ian M. Banks wrote of this. How the machine minds could comfortably write opera's greater than any man, but still humans would go to the theatre, just to appreciate it, but the impact was recognised in society. Of course that world was based on post-scarcity whilst we are not.

Re: Sora: Creating video from text

#942
post #873

This is both amazing and saddening to me. All our cultural legacy is being fed into a monstrous machine that gives no attribution to the original content with which it was fed, and so the creative industry seems to be in great danger. Creativity being automated while humans are forced to perform menial tasks for minimum wage doesn't seem like a great future and the geriatric political class has absolutely no clue how…

Machines can reliably beat humans at chess. Has that stopped anyone from playing? Has it stopped anyone from watching chess tournaments?

Re: Sora: Creating video from text

#943

This is all very impressive. I can't help to wonder though. How is text-to-video going to benefit humanity? That's what OpenAI is supposedly about, right? We'll get some groundbreaking film content out of this in the hands of a few talented creatives, and a vast ocean of mediocre content from the hands of talentless people who know how to type. What's the benefit to humanity, concretely?

How else is the next generation of talented creatives cultivated, if not out of the pool of the millions of untalented typists?

Re: Sora: Creating video from text

#944

I think the implications go much further than just the image/video considerations. This model shows a very good (albeit not perfect) understanding of the physics of objects and relationships between them. The announcement mentions this several times. The OpenAI blog post lists "Archeologists discover a generic plastic chair in the desert, excavating and dusting it with great care." as one of the "failed" cases. But t…

> In the Tokyo one, the model is smart enough to figure out that on a train, the reflection would be of a passenger, and the passenger has Asian traits since this is Tokyo.

How is this any more accurate than saying that the model has mostly seen Asian people in footage of Tokyo, and thus it is most likely to generate Asian-features for a video labelled "Tokyo"? Similarly, how many videos looking out a train window do you think it's seen where there was not a reflection of a person in the window when it's dark?

Re: Sora: Creating video from text

#945
post #623

I think the implications go much further than just the image/video considerations. This model shows a very good (albeit not perfect) understanding of the physics of objects and relationships between them. The announcement mentions this several times. The OpenAI blog post lists "Archeologists discover a generic plastic chair in the desert, excavating and dusting it with great care." as one of the "failed" cases. But t…

I found the one about the people in Lagos pretty funny. The camera does about a 360deg spin in total, in the beginning there are markets, then suddenly there are skyscrapers in the background. So there's only very limited object permanence. > A beautiful homemade video showing the people of Lagos, Nigeria in the year 2056. Shot with a mobile phone camera. > https://cdn.openai.com/sora/videos/lagos.mp4

> then suddenly there are skyscrapers in the background. So there's only very limited object permanence.

Ah but you see that is artistic liberty. The director wanted it shot that way.

Re: Sora: Creating video from text

#946

I'd love to feel excited by all these advancements and somehow I feel numb. I get part of the feeling (worry about inequalities it may generate), but I sense something more. It's like I see it as a toy... I'm unable to dream on how this will impact my life in any meaningful way.

Imagine dumping all the HIPAA data into a process like this. Obviously fraught with privacy and accuracy[0] concerns. Nonetheless, it might help us move some things forward.

Re: Sora: Creating video from text

#947
post #404

This is all very impressive. I can't help to wonder though. How is text-to-video going to benefit humanity? That's what OpenAI is supposedly about, right? We'll get some groundbreaking film content out of this in the hands of a few talented creatives, and a vast ocean of mediocre content from the hands of talentless people who know how to type. What's the benefit to humanity, concretely?

> Sora serves as a foundation for models that can understand and simulate the real world, a capability we believe will be an important milestone for achieving AGI.

That struck me as a line they added to drum up more funding.

Re: Sora: Creating video from text

#949

Call me whatever you want, but this technology should not exist. People to just create lifelike videos of anything they can put their mind to, is bound to lead to the ruining of many peoples' lives. As many people that are aware and interested in this technology, there is 100x people who have no idea, don't care or can't comprehend it. Those are the people that I fear for. Grab a few pictures of the grandkids off of…

"Writing should be restricted to the educated few who can responsibly carry the Church's message." -anti-technologist from 1000 years ago, probably

Re: Sora: Creating video from text

#950
post #96

My AI idea: Civil war as a service (CWaaS) Prompt: poll worker sneakily taking ballots labeled , and throwing them in the trash.

You realize how easy it is to do that with actors, right?

At least 1000x more effort than typing a sentence into your keyboard. Hence less likely to happen at the same frequency and scale.
Post reply on HN