Live data from Hacker News

Sora is here

openai.com

351–360 of 1001 posts

Re: Sora is here

#351
post #281

Earlier quoted context omitted.

Take the example to the extreme: In 10 years, I prompt my photo album app "Generate photorealistic video of my mother playing with a ladybug". The juxtaposition of something that looks extremely real (your mother) and something that never happened (ladybug) is something that's hard for the mind to reconcile. The presence of a real thing inadvertently and subconsciously gives confidence to the fake thing also being re…

I think this hooks in quite well to the existing dialogue about movies in particular. Take an action movie. It looks real but is entirely fabricated. It is indeed something that society has to shift to deal with. Personally, I'm not sure that it's the photoreal aspect that poses the biggest challenge. I think that we are mentally prepared to handle that as long as it's not out of control (malicious deep-fakes used to…

I looked at the Sora videos and all the subject "weights" and "heft" are off. And in the same way that Anna Taylor-Joy's jump in the The Gorge at the end of the new movie trailer looked not much better than years-ago Spiderman swinging on a rope.

Re: Sora is here

#352
post #281

Earlier quoted context omitted.

Take the example to the extreme: In 10 years, I prompt my photo album app "Generate photorealistic video of my mother playing with a ladybug". The juxtaposition of something that looks extremely real (your mother) and something that never happened (ladybug) is something that's hard for the mind to reconcile. The presence of a real thing inadvertently and subconsciously gives confidence to the fake thing also being re…

I think this hooks in quite well to the existing dialogue about movies in particular. Take an action movie. It looks real but is entirely fabricated. It is indeed something that society has to shift to deal with. Personally, I'm not sure that it's the photoreal aspect that poses the biggest challenge. I think that we are mentally prepared to handle that as long as it's not out of control (malicious deep-fakes used to…

I'm still waiting on the future waves of PTSD from hyper realistic horror games. I can't think of a worse thing to do then hand a kid a VR headset (or game system) and have them play a game that is designed to activate every single fight or flight nerve in the body on a level that is almost indistinguishable from reality. 20 years ago that would have been the plot to a torture porn flick.

Even worse than that is when people get USED to it and no longer have a natural aversion to horrific scenes taking place in the real world.

This AI stuff accelerates that process of illusion but in every possible direction at once.

As much as people don't want to believe it, by beholding we are indeed changed.

Re: Sora is here

#353
post #294
post #245

Earlier quoted context omitted.

I don't understand why you see a distinction between models that generate text, and those that generate images, video or audio. They're all digital formats, and the technology itself is fairly agnostic about what it's actually generating. Can't text also be considered art? There's as much art in poetry, lyrics, novels, scripts, etc. as in other forms of media. The thing is that the generative tech is out of the bag,…

Simple: I am equally offput when LLMs are used for generating poetry, lyrics, novels, scripts, etc. I don't like it when low-effort generated slop is passed off as art . I just think that LLMs have genuine use for non-artistic things, which is why I said it's dangerous but may be useful if we play our cards right.

the offensive part is that it's creative theft by digesting other people's creative works then reworked and regurgitated. It's 'fine' when it's technical documentation and reference work, but that's not human expression.

Re: Sora is here

#354

A little worried how young children watching these videos may develop inaccurate impressions of physics in nature. For instance, that ladybug looks pretty natural, but there's a little glitch in there that an unwitting observer, who's never seen a ladybug move before, may mistake as being normal. And maybe it is! And maybe it isn't? The sailing ship - are those water movements correct? The sinking of the elephant int…

Yes Bugs bunny and willie the coyote harmed ours physics.

Re: Sora is here

#355

A little worried how young children watching these videos may develop inaccurate impressions of physics in nature. For instance, that ladybug looks pretty natural, but there's a little glitch in there that an unwitting observer, who's never seen a ladybug move before, may mistake as being normal. And maybe it is! And maybe it isn't? The sailing ship - are those water movements correct? The sinking of the elephant int…

YouTube Shorts are full of AI animal videos with distorted proportions, living in the wrong habitat, and so on. They popped up on my son’s account and I hate them for the reasons you outline. They aren’t cartoonish enough explain away, nor realistic enough to be educational.

And have you watched the brain rot that is Tik toks?

Re: Sora is here

#356
post #147

Earlier quoted context omitted.

> Now expand that to movies and games and you can get why this whole generative-AI bubble is going to pop. What will save it is that, no matter how picky you are as a creator, your audience will never know what exactly was that you dreamed up, so any half-decent approximation will work. In other words, a corollary to your corollary is, "Fortunately, you don't need them to be, because no one cares about low-order bits…

Your eye sees just about every frame of a film… People may not think they care, but obviously they do. That’s why marvel movies do better than DC ones. People absolutely care about details in their media.

Fair point, particularly given the example. My conclusion wrt. Marvel vs. DC is that DC productions care much less about details, in exactly the way I find off-putting.

Not all details matter, some do. And, it's better to not show the details at all, than to be inconsistent in them.

Like, idk., don't identify a bomb as a specific type of existing air-fuel ordnance and then act about it as if it was a goddamn tactical nuke. Something along these lines was what made me stop watching Arrow series.

Re: Sora is here

#357

A little worried how young children watching these videos may develop inaccurate impressions of physics in nature. For instance, that ladybug looks pretty natural, but there's a little glitch in there that an unwitting observer, who's never seen a ladybug move before, may mistake as being normal. And maybe it is! And maybe it isn't? The sailing ship - are those water movements correct? The sinking of the elephant int…

> A little worried how young children watching these videos may develop inaccurate impressions of physics in nature.

And why don't we worry this about CGI?

CGI is not always made with a full physical simulation, and is not always intended to accurately represent real-world physics.

Re: Sora is here

#358
post #220
post #164

I got lucky and got in moments after it launched, managed to get a video of "A pelican riding a bicycle along a coastal path overlooking a harbor" and then the queue times jumped up (my second video has been in the queue for 20+ minutes already) and the https://sora.com site now says "account creation currently unavailable" Here's my pelican video: https://simonwillison.net/2024/Dec/9/sora/

I don't have a lot of mental model for how this works, but I was surprised to note that it seems to maintain continuity on the shapes of the bushes and brown spots on the grass that track out of frame on the left and then reappear as it pans back into frame.

[deleted]

Re: Sora is here

#359

Wow this is bad. And by bad i mean worse than leading open source and existing alternatives. Is it me or does it seem like OpenAI revolutionized with both chatGPT and Sora, but they've completely hit the ceiling? Honestly a bit surprised it happened so fast!

Could it be that text sources are plenty, and more dense than training for videos, and images?

Re: Sora is here

#360
post #149

Earlier quoted context omitted.

Those are indeed 3 papers.

Yes in a nutshell they explain that you can express a picture or a video with relatively few discrete information. First paper is the most famous and prompted a lot of research to using text generation tools in the image generation domain : 256 "words" for an image, Second paper is 24 reference image per minutes of video, Third paper is a refinement of the first saying you only need 32 "tokens". I'll let you multiply…

I think something is getting lost in translation.

These papers, from my quick skim (tho I did read the first one fully years ago,) seem to show that some images and to an extent video can be generated from discrete tokens, but does not show that exact images nor that any image can be.

For instance, what combination of tokens must I put in to get _exactly_ Mona Lisa or starry night? (Tho these might be very well represented in the data set. Maybe a lesser known image would be a better example)

As I understand, OC was saying that they can’t produce what they want with any degree of precision since there’s no way to encode that information in discrete tokens.

Post reply on HN