Live data from Hacker News

Make-A-Video: AI system that generates videos from text

makeavideo.studio

241–250 of 399 posts

Re: Make-A-Video: AI system that generates videos from text

#241

Earlier quoted context omitted.

AI is eating the world, and the vast majority of people are not paying attention. I don't know what artists, truck drivers, Uber/Lyft/taxi drivers, delivery drivers, programmers, doctors, judges, fast food workers, etc. are going to do.

doctors ? We are a looooooong way to go before AI can replace actual doctors.

A month ago I would have said that about actual artists.

Re: Make-A-Video: AI system that generates videos from text

#242
post #179

Earlier quoted context omitted.

I love this! But after trying it a few times I got this result :). So fascinating. https://thismoviedoesnotexist.org/movie/the-terminator Brings up the age-old question of how much the learning in these models is just memorization. Though in cases like these it’s hard to tell.

Not sure how I feel about an AI generating a movie concept that involves a "rise of the machines".

I feel like "generate" is kind of a strong word in this case though. At this rate if the machines rise up, they will do so just to parrot all the "machines rise up" plot synopses in their training corpus.

Re: Make-A-Video: AI system that generates videos from text

#243

As an owner of a Video Production studio, this kind of tech is blowing my mind and makes me equally excited and scared. I can see how we could incorporate such tools in our workflows, and at the same time I'm worried it'll be used to spam the internet with thousands and thousands of souless generated videos, making it even harder to look through the noise. A fun related experiment, I thought it was fun to see what ki…

No post body was provided.

Re: Make-A-Video: AI system that generates videos from text

#244
post #96
post #8

Earlier quoted context omitted.

In case anyone else had problems finding it - it's at the bottom of the page: https://makeavideo.studio/assets/a_golden_retriever_eating_i... (webp) That grasp though. These things still feel a bit like e.g. Google/GCP services to me: Super appealing at first glance, quite close to what you want, but somehow never quite there. Maybe they'll asymptotically get there, eventually? Perhaps that statistical model can't re…

It may be that it's the deep learning tech which will never quite get there. GPT-3 has similar shortcomings in its mimicry. We're 95% there, I guess, but may never quite reach 100%.

Nah, the current issues are just because we're trying to do everything in one step. Because we've built tools that have so much of a stimulus-response approach, few efforts have been made toward interfaces that ask for clarification ('when you say X, do you mean XYZ or XXX?').

Image-to-image and tuning already addresses many of these issues; just as inpainting works really well, it won't be long before we have select-and-repair, where you add an additional prompt like 'improve this part - the ice cream is fine, just work on the dog's muzzle.'

Re: Make-A-Video: AI system that generates videos from text

#245
Clearly, attention really is all you need.

Are GPU vendors (well, gpu vendor, as far as I can tell) focusing heavily on increasing VRAM? My understanding is that transformers are pretty quick to train, but have significant memory costs.

When they say that video is infeasible with memory... does that mean that if we had enough memory (128? 256? gb) we would be able to realistically train such networks with temporal attention?

This is insanely exciting. It looks like we are limited, at this point, only by compute.

Re: Make-A-Video: AI system that generates videos from text

#246

Earlier quoted context omitted.

Make blockbusters, because scenes that would be incredibly expensive to shoot now will be practically free with this?

Every 15 year old kid will be making their own blockbusters

Have you talked to the average 15 year old kid? Hell, have you talked to the average 50 year old? It will be the same as ever, a sea of absolute shit surrounding some true gems, be they from novel creativity or just excellent execution of well worn ground. The role of the trusted curator will rise and brands will gain more power.

I am sure I will watch MY 15 year old's attempts, and maybe a few from my extended circle but most of my consumed content will still come from what makes the cut to Netflix or HBO etc. Technology like this will empower the truly creatives once it has matured. I would expect closer to 20 years than 5 however.

Re: Make-A-Video: AI system that generates videos from text

#247
post #115

Earlier quoted context omitted.

They forgot ‘trending on onlyfans’ But actually, this technology is super exciting. Imagine a future where movies and games are choose your own adventure.

We'll have procedural generation that will be hard to distinguish from human-made content. Goodbye repetitive Skyrim filler caves!

I've never played Skyrim - was that the immersive 3d version of a maze of twisty little passages?

Besides just textural content, it's intriguing to consider the possibilities of full-3d roguelikes.

Re: Make-A-Video: AI system that generates videos from text

#248

What's mind blowing is that you can extrapolate where this is going to go. Eventually, you will be able to generate full movie scenes from descriptions. What's interesting to me is how this is so similar to human imagination. Give me a description and I will fabricate the visuals in my mind. Some aspects will be detailed, others will be vague, or unimportant. Crazy to see how fast AI is progressing. Machines are appr…

Related: https://pub.towardsai.net/stable-diffusion-based-image-compr...

Re: Make-A-Video: AI system that generates videos from text

#249

I really don’t understand the fear people have about these things. have I missed something and everyone else was placing huge value in out of context videos and pictures? if you read “France declares war on Canada”, you’re not gonna believe it unless it’s coming from an extremely reputable source. so why would you trust a random unsourced video? the absolute worst thing that’s gonna happen is that video-based social…

We are quickly approaching a point where these independently created AI systems will be wrangled together the same away protocols were to create a computer network and the AI that emerges will likely be able to create its own code, solve its own problems and generate its own recursive processes. Everything at a conscious level without the ability to self recognize. If AI ever does wake up; we'll never know - first thing that happens is it will hide from us.

The fear is real and only seems fantastical because life is often stranger than fiction.

Instead of training on data these AIs will soon train on "creativity" and these layers of containerized thought will merge.

Re: Make-A-Video: AI system that generates videos from text

#250

As an owner of a Video Production studio, this kind of tech is blowing my mind and makes me equally excited and scared. I can see how we could incorporate such tools in our workflows, and at the same time I'm worried it'll be used to spam the internet with thousands and thousands of souless generated videos, making it even harder to look through the noise. A fun related experiment, I thought it was fun to see what ki…

For those who are scared about this technology, it’s good to look at what AI has done to Chess. The best chess seems to be when AI is used along with humans. I think image and video AI will best be exploited when human input is also taken into account. There is still something special about human creativity, I think AI will just be another tool to expand that. At least, in the short term I would say 10 years perhaps.…

I think a key difference here is that with chess, 'goodness' is defined by winning. With content generation, the training methods point towards some form of comparing the generated thing to some observed data, but the 'goodness' of the content from the perspective of potentially competing with or displacing human creators is "do people like to consume it?"

If one trained using e.g. a tiktok like dataset showing viewer response measurements for each video, and do conditional generation on those response values ("prank video watchers are highly likely to watch the full video"), are we really that far from a system that learns to generate content that attracts and hold eyeballs? Not so long ago there were a lot of concerning trend pieces about how youtube had a network of creators making bizarre, disturbing or transfixing videos being watched entirely by young children. Before that, it was clickbait listicles. "Bad" content that can get eyeballs can still wildly steer what humans create and consume. I'm wondering if in 2 years we'll have an enormous number of short videos that we all agree are "bad" but which are nevertheless constantly watched.

Post reply on HN