Live data from Hacker News

Sora 2

openai.com

281–290 of 916 posts

Re: Sora 2

#281
post #186

Earlier quoted context omitted.

Water level in a glass changing between shots is one thing, the protagonist’s face and clothes changing is another.

Well put. Honestly the actor part is mostly solved by now, the tricky part is depicting any kind of believable, persistent space across different shots. Based off of amateur outputs from places like https://www.reddit.com/r/aivideo/ , at least! This release is clearly capable of generating mind-blowingly realistic short clips, but I don't see any evidence that longer, multi-shot videos can be automated yet. With a pr…

[deleted]

Re: Sora 2

#282
post #98
post #20

Earlier quoted context omitted.

one of the example prompts is literally: Prompt: in the style of a studio ghibli anime, a boy and his dog run up a grassy scenic mountain with gorgeous clouds, overlooking a village in the distant background

Wow that is dark, after Ghiblis staunch stance on AI. These companies and their shareholders really are complete scum in my eyes, just like AI in miltech. Not because the tech isn't super interesting but because they steal years of hard work and pain from actual artists with zero compensation - and then they brag about it in the most horrible way possible, with zero empathy. Then comes losing the little humanity left…

Indeed is difficult to NOT share this resentment, should anyone understand what actually happens.

Re: Sora 2

#283

Earlier quoted context omitted.

Thing that was previously very expensive, manual and took a long time to do, and is done A LOT, is now made faster and cheaper by computers. Pretty much the same problem we all work on every day in $DAY_JOB.

Fun and games until someone uses a tool like this to scam your family

It's ok, they're making the market for anti-ai tools much much bigger. (whether those tools work or not is a different issue)

Re: Sora 2

#284
post #117

The main lesson I learned from the March ChatGPT image generation launch - which signed up 100 million new users in the first week - is that people love being able to generate images of their friends and family (and pets). I expect the "cameo" feature is an attempt at capturing that viral magic a second time.

[flagged]

True, but our old technology like radios has been doing that for a long time too.

Re: Sora 2

#285

Earlier quoted context omitted.

I actually wonder if this will kill off the social apps and the bragging that happens. It will be flooded by people faking themselves doing the unimaginable.

This is also my thesis. The internet is going to be saturated with AI slop indiscernible from real content. Once it reaches a tipping point, there will no longer be much of a reason to consume the content at all. I think social networks that can authenticate video/photo/text content as human-created will be a major trend in a few years.

I have no clue if the reactions are real, but there are some videos online of people showing their grandparents gameplay from Grand Theft Auto games trying to convince them that it is real footage. The point of the videos is to laugh at their reactions where they question if it really happened, etc.

Maybe this will result in something similar, but it can affect more people who aren’t as wary.

Re: Sora 2

#286
post #198

Going to be an amazing source of training data, wait till they get it to real time and people are leaving their video camera open for AR features. OpenAI is about to have a lot of current real world image data, never mind the sentiment analysis.

I don't think they were limited for video training data. Gathering real world data is pretty easy, gathering curated information is a little more difficult.

Re: Sora 2

#287

I'm a software engineer and hobbyist actor/director. My friends are in the film industry and are in IATSE and SAG-AFTRA. I've made photons-on-glass films for decades, and I frequently film stuff with my friends for festivals. I love this AI video technology. Here are some of the films my friends and I have been making with AI. These are not "prompted", but instead use a lot of hand animation, rotoscoping, and human v…

Well I was entertained.

What is up with a lot of voices are left ear only?

Re: Sora 2

#288

Impressively high level of continuity. The only errors I could really call out are: 1/ 0m23s: The moon polo players begin with the red coat rider putting on a pair of gloves, but they are not wearing gloves in the left-vs-right charge-down. 2/ 1m05s: The dragon flies up the coast with the cliffs on one side, but then the close-up has the direction of flight reversed. Also, the person speaking seemingly has their back…

Not sure if it counts as a continuity error, but in the example "Prompt: Martial artist doing a bo-staff kata waist-deep in a koi pond", his wooden staff changes shape several times, resembling a bow at points. That was the first example I noticed as "clearly AI."

Re: Sora 2

#289

Impressively high level of continuity. The only errors I could really call out are: 1/ 0m23s: The moon polo players begin with the red coat rider putting on a pair of gloves, but they are not wearing gloves in the left-vs-right charge-down. 2/ 1m05s: The dragon flies up the coast with the cliffs on one side, but then the close-up has the direction of flight reversed. Also, the person speaking seemingly has their back…

The Bo staff in the koi pond also seems to involve some impossible wrist movements

Re: Sora 2

#290
post #215

Earlier quoted context omitted.

Is there a reason that's superior to subtitles, which are already fairly easy to generate?

sign languages are completely different languages from spoken languages, with their own grammar etc. subtitles can work but it's basically a second language. perhaps comparable to many countries where people speak a dialect that's very different from the "standard" written language. this is why you sometimes have sign language interpreters at events, rather than just captions. there's not really a widely accepted wri…

>this is why you sometimes have sign language interpreters at events, rather than just captions.

No, the reason is because a) it's in real time, and b) there's no screen to put the subtitles on. If it was possible to simply display subtitles on people's vision, that would be much more preferable, because writing is a form of communication more people are familiar with than sign language. For example, someone might not be deaf, but might still not be able to hear the audio, so a sign language interpreter would not help them at all, while closed captions would.

Post reply on HN