Live data from Hacker News

Veo 2: Our video generation model

deepmind.google

51–60 of 342 posts

Re: Veo 2: Our video generation model

#51
post #27
post #6

This looks great, but I'm confused by this part: > Veo sample duration is 8s, VideoGen’s sample duration is 10s, and other models' durations are 5s. We show the full video duration to raters. Could the positive result for Veo 2 mean the raters like longer videos? Why not trim Veo 2's output to 5s for a better controlled test? I'm not surprised this isn't open to the public by Google yet, there's a huge amount of volu…

> I'm not surprised this isn't open to the public by Google yet, Closed models aren't going to matter in the long run. Hunyuan and LTX both run on consumer hardware and produce videos similar in quality to Sora Turbo, yet you can train them and prompt them on anything. They fit into the open source ecosystem which makes building plugins and controls super easy. Video is going to play out in a way that resembles image…

Stable Diffusion and Flux did not win though. Midjourney and chatGPT won.

Re: Veo 2: Our video generation model

#52
post #48

OpenAI is like the super luxurious yacht all pretty and shiny, while Google's AI department is the humongous nuclear submarine at least 5 times bigger than the yacht with a relatively cool conning tower, but not that spectacular to look at. Like the tanker which is still steering to fully align with the course people expect it to be, which they don't recognize that it will soon be there and be capable of rolling over…

google definitely does not have AGI hhaaha

Re: Veo 2: Our video generation model

#53
post #43

Earlier quoted context omitted.

Put another way, over time people devalue things which can be produced with minimal human effort. I suspect it's less about humanity's values, and more about the way money closely tracks "time" (specifically the duration of human effort).

I strongly disagree. How many clothes do you buy that have 100 thread count, and are machine-made, vs hand-knit sweaters or something? When did you ask people for directions, or other major questions, instead of Google? You can wax poetic about wanting "the human touch", but at the end of the day, the market speaks -- people will just prefer everything automated. Including their partners, after your boyfriend can rem…

> PS: anything you write on HN can already have been written by AI, pretty soon you may as well quit producing any content at all. No one will care whether you wrote it.

People theoretically would care, but the internet has already set up producing things to be pseudo-anonymous, so we have forgotten the value of actually having a human being behind content. That's why AI is so successful, and it's a damn shame.

Re: Veo 2: Our video generation model

#54
post #4
post #2

Judging by how they've been trying to ram AI into YouTube creators workflows I suppose it's only a matter of time before they try to automate the entire pipeline from idea, to execution, to "engaging" with viewers. It won't be good at doing any of that but when did that ever stop them. https://www.youtube.com/watch?v=26QHXElgrl8 https://x.com/surri01/status/1867433782992879617

And then suddenly this is not something that fascinates people anymore… in 10 years as non-synthetic becomes the new bio or artisan or whatever you like. Humanity has its ways of objecting accelerationism.

> Humanity has its ways of objecting accelerationism.

Actually, typically human objection only slows it down and often it becomes a fringe movement, while the masses continue to consume the lowest common denominator. Take the revival of the flip phone, typewriter, etc. Sadly, technology marches on and life gets worse.

Re: Veo 2: Our video generation model

#55

I appreciate they posted the skateboarding video. Wildly unrealistic whenever he performs a trick - just morphing body parts. Some of the videos look incredibly believable though.

It is great so see a limitations section. What would be even more honest is a very large list of videos generated without any cherry picking to judge the expected quality for the average user. Anyway, the lack of more videos suggests that there might be something wrong somewhere.

Re: Veo 2: Our video generation model

#57

This might be a dumb question to ask, but what exactly is this useful for? B-Roll for YouTube videos? I'm not sure why so much effort is being put into something like this when the applications are so limited.

Are they that limited? It's a machine that can make videos from user input: it can ostensibly be used wherever you need video, including for creative, technical and professional applications.

Now, it may not be the best fit for those yet due to its limitations, but you've gotta walk before you can run: compare Stable Diffusion 1.x to FLUX.1 with ControlNet to see where quality and controllability could head in the future.

Re: Veo 2: Our video generation model

#59

This might be a dumb question to ask, but what exactly is this useful for? B-Roll for YouTube videos? I'm not sure why so much effort is being put into something like this when the applications are so limited.

Back when computers took up a whole room, you'd also have asked: "but what exactly is this useful for? B-Roll some simple calculations that anybody can do with a piece of paper and a pen."?

Think 5-10 years into the future, this is a stepping stone

Re: Veo 2: Our video generation model

#60
post #18

Last time Google made a big Gemini announcement, OpenAI owned them by dropping the Sora preview shortly after. This feels like a bit of a comeback as Veo 2 (subjectively) appears to be a step up from what Sora is currently able to achieve.

[deleted]
Post reply on HN