Live data from Hacker News

Make-A-Video: AI system that generates videos from text

makeavideo.studio

1–10 of 399 posts

Re: Make-A-Video: AI system that generates videos from text

#3
This looks like the video equivalent of Dall-E 1. Hard to believe how far we've come so quickly.

The paper talks about "pseudo 3D attention layers" that are used in place of temporal attention layers for each dimension due to memory consumption. It seems like AI research is vastly outpacing GPU development.

Re: Make-A-Video: AI system that generates videos from text

#4
I'm rooting for this tech. Hopefully this will get modern movies out of their low risk reboot loop since it will be cheaper to make a movie that have new story lines that are commercially untested. I'd be happy to watch a movie that doesn't look AAA, but has compelling writing and makes me think. Or maybe I'll just stick to books.

Re: Make-A-Video: AI system that generates videos from text

#7

The > A golden retriever eating ice cream on a beautiful tropical beach at sunset, high resolution example is terrifying.

They forgot ‘trending on onlyfans’

But actually, this technology is super exciting. Imagine a future where movies and games are choose your own adventure.

Re: Make-A-Video: AI system that generates videos from text

#8

The > A golden retriever eating ice cream on a beautiful tropical beach at sunset, high resolution example is terrifying.

In case anyone else had problems finding it - it's at the bottom of the page:

https://makeavideo.studio/assets/a_golden_retriever_eating_i... (webp)

That grasp though.

These things still feel a bit like e.g. Google/GCP services to me: Super appealing at first glance, quite close to what you want, but somehow never quite there. Maybe they'll asymptotically get there, eventually? Perhaps that statistical model can't really make it to the level we want it to?

Re: Make-A-Video: AI system that generates videos from text

#9
post #3

This looks like the video equivalent of Dall-E 1. Hard to believe how far we've come so quickly. The paper talks about "pseudo 3D attention layers" that are used in place of temporal attention layers for each dimension due to memory consumption. It seems like AI research is vastly outpacing GPU development.

Looks a bit better than DALLE1 IMHO. They've demonstrated greater range.
Post reply on HN