Live data from Hacker News

Veo 2: Our video generation model

deepmind.google

81–90 of 342 posts

Re: Veo 2: Our video generation model

#82

This might be a dumb question to ask, but what exactly is this useful for? B-Roll for YouTube videos? I'm not sure why so much effort is being put into something like this when the applications are so limited.

We're preparing to use video generation (specifically image+text => video so we can also include an initial screenshot of the current game state for style control) for generating in-game cutscenes at our video game studio. Specifically, we're generating them at play-time in a sandbox-like game where the game plays differently each time, and therefore we don't want to prerecord any cutscenes.

Re: Veo 2: Our video generation model

#83

Earlier quoted context omitted.

> PS: anything you write on HN can already have been written by AI, pretty soon you may as well quit producing any content at all. No one will care whether you wrote it. People theoretically would care, but the internet has already set up producing things to be pseudo-anonymous, so we have forgotten the value of actually having a human being behind content. That's why AI is so successful, and it's a damn shame.

What exactly is the value of having a human behind content if it gets to the point that content generated by AI is indistinguishable from content generated by humans?

I think "indistinguishable" is a receding horizon. People are already good at picking out AI text, and AI video is even easier. Even if it looks 100% realistic on the surface, the content itself (writing, concept, etc) will have a kind of indescribable "sameness" that will give it away.

If there's one thing that connects all media made in human history, it's that humans find humans interesting. No technology (like literally no technology ever) will change that.

Re: Veo 2: Our video generation model

#84
post #69

Earlier quoted context omitted.

> Humanity has its ways of objecting accelerationism. Actually, typically human objection only slows it down and often it becomes a fringe movement, while the masses continue to consume the lowest common denominator. Take the revival of the flip phone, typewriter, etc. Sadly, technology marches on and life gets worse.

Does life get worse for the majority of people or do the fruits of new technology rarely address any individual person’s progress toward senescence? (The latter feels like tech moves forward but life gets worse.)

Of course, it depends on how you define "worse". If you use life expectancy, infant mortality, and disease, then life has in the past gotten better (although the technology of the past 20 years has RARELY contributed to any of that).

If you use 'proximity to wild nature', 'clean air', 'more space', then life has gotten worse.

But people don't choose between these two. They choose between alternatives that give them analgesics in an already corrupt society creating a series of descending local maximae.

Re: Veo 2: Our video generation model

#85

Earlier quoted context omitted.

> PS: anything you write on HN can already have been written by AI, pretty soon you may as well quit producing any content at all. No one will care whether you wrote it. People theoretically would care, but the internet has already set up producing things to be pseudo-anonymous, so we have forgotten the value of actually having a human being behind content. That's why AI is so successful, and it's a damn shame.

What exactly is the value of having a human behind content if it gets to the point that content generated by AI is indistinguishable from content generated by humans?

The fact that anyone would ask this question is incredible!

It's so we can in a fraction of those cases, develop real relationships to others behind the content! The whole point of sharing is to develop connections with real people. If all you want to do is consume independently of that, you are effectively a soulless machine.

Re: Veo 2: Our video generation model

#86

This might be a dumb question to ask, but what exactly is this useful for? B-Roll for YouTube videos? I'm not sure why so much effort is being put into something like this when the applications are so limited.

in it's current state, it's already useful for b-roll, video backgrounds for websites, and any other sort of "generic" application where the point of the shot is just to establish mood and fill time.

but more than anything it's useful as a stepping stone to more full-featured video generation that can maintain characters and story across multiple scenes. it seems clear that at some point tools like this will be able to generate full videos, not just shots.

Re: Veo 2: Our video generation model

#87
post #65

Earlier quoted context omitted.

Does everyone have "legal" access to YouTube. In theory that should matter to something like Open(Closed)Ai. But who knows.

I mean, I have trained myself on Youtube. Why can't a silicon being train itself on Youtube as well?

Because silicon is a robot. A camcorder can't catch a flick with me in the theater even if I dress it up like a muppet.

Re: Veo 2: Our video generation model

#88
post #27

Earlier quoted context omitted.

> I'm not surprised this isn't open to the public by Google yet, Closed models aren't going to matter in the long run. Hunyuan and LTX both run on consumer hardware and produce videos similar in quality to Sora Turbo, yet you can train them and prompt them on anything. They fit into the open source ecosystem which makes building plugins and controls super easy. Video is going to play out in a way that resembles image…

Stable Diffusion and Flux did not win though. Midjourney and chatGPT won.

“Won” what exactly? I have no issues running stable diffusion locally.

Since Llama3.3 came out it is my first stop for coding questions, and I’m only using closed models when llama3.3 has trouble.

I think it’s fairly clear that between open weights and LLMs plateauing, the game will be who can build what on top of largely equivalent base models.

Re: Veo 2: Our video generation model

#89

Earlier quoted context omitted.

google definitely does not have AGI hhaaha

Yeah pretty bad example from parent but the point stands I think... I mostly just assume that for everything ChatGPT hypes/teases Google probably has something equivalent internally that they just aren't showing off to the public.

I know that Google's internal ChatGPT alternative was significally worse than ChatGPT(confirmed both in news and by Googlers) around a year back. So you might say they might overtake OpenAI because of more resources, but they aren't significantly ahead of OpenAI.

Re: Veo 2: Our video generation model

#90
post #31

Earlier quoted context omitted.

Everyone has access to YouTube. It’s safe to assume that Sora was trained on it as well.

All you can eat? Surely they charge a lot for that, at least. And how would you even find all the videos?

They already did it, and I’m guessing they were using some of the various YouTube down loaders Google has been going after.
Post reply on HN