Live data from Hacker News

Ovi: Twin backbone cross-modal fusion for audio-video generation

github.com

101–110 of 122 posts

Re: Ovi: Twin backbone cross-modal fusion for audio-video generation

#101

Earlier quoted context omitted.

you're seeing them on reddit and twitter. they aren't normal people. normal people don't run a background check on $thing and its creators to determine whether they are allowed to like it or not.

In another comment you're defending AI girlfriends, and you're going to tell me what "normal" people who don't live on the internet do? As a matter of fact, all the actually normal people I talk to about AI in person also find it offputting.

that comment doesn't "defend" AI girlfriends, it points out the absurdity of condemning AI girlfriends in a world where porn and prostitution exist and are widely accepted by the "polite society".

also, case in point, normal people don't dig through a random stranger's post history to look for an ad hominem opportunity, and instead evaluate individual posts by their contents. lol.

Re: Ovi: Twin backbone cross-modal fusion for audio-video generation

#102
post #96

Earlier quoted context omitted.

Just like with games visual are only part of the formula. In theory you can make a truly fantastic movie (or game) for next to nothing. I didnt believe this before The man from earth. That cost 200K but it didnt have to. Edit: perhaps 12 angry men was good enough at the time.

>Edit: perhaps 12 angry men was good enough at the time. I recently watched it for the first time, and it was one of the best movies I've seen. I can't believe how invested I was, even though the plot was so simple.

Because the movie is the epitome of humanism.

When AI slop figures out that formula, we are truly cooked

Re: Ovi: Twin backbone cross-modal fusion for audio-video generation

#103

Earlier quoted context omitted.

In another comment you're defending AI girlfriends, and you're going to tell me what "normal" people who don't live on the internet do? As a matter of fact, all the actually normal people I talk to about AI in person also find it offputting.

that comment doesn't "defend" AI girlfriends, it points out the absurdity of condemning AI girlfriends in a world where porn and prostitution exist and are widely accepted by the "polite society". also, case in point, normal people don't dig through a random stranger's post history to look for an ad hominem opportunity, and instead evaluate individual posts by their contents. lol.

The combination of a brand new account with an incredibly juvenile username and careless writing (lack of capitalization in your case) usually is a red flag for a spam account, so yes, in these cases I usually check comments to see if I'm wasting my time with a troll.

Porn is still taboo. It's understood that most people use it, but it's not exactly something you bring up in polite company.

Where on earth do you live that prostitution is "widely accepted by polite society"? You can go to jail for it where I am.

And I did address the rest of your comment. As I said, in my experience "normal" people do object to AI content. I don't know where you got the bit about "background checks" and being "allowed" to like stuff. Nobody I know had to be told to have an aversion to AI "art", it's a natural reaction.

Re: Ovi: Twin backbone cross-modal fusion for audio-video generation

#104

Earlier quoted context omitted.

I'm vaguely reminded of the excellent Jackbox game Tee Fury, in which players submit slogans for T shirts and "art" separately. Players then get to choose from a few options for slogans and designs to make T shirts which are voted on by the group. I have fond memories of laughing until I was in tears when playing with a group of friends over drinks during the lockdowns in 2020. Something about the process just natura…

T shirt game is the best jackbox game! Whenever one of my friend groups is gathered we always make it a point to do an exquisite corpse story on a piece of paper while we’re inebriated in some way xD Video version will be wild

It's seriously so good, in fact it's so good that every other Jackbox game is vaguely disappointing because nothing is half as fun as Tee Fury lol.

Re: Ovi: Twin backbone cross-modal fusion for audio-video generation

#105
post #21
post #10

Earlier quoted context omitted.

Soon as the video models can keep characters consistent across scenes. It could take months of prompting to get each scene, but regular movies have long shooting timelines too. If we ever get to instant movie, thought to scene, then movies will die since people will just daydream through the AI.

I can see soap-opera-style, video-manga becoming a thing, where you get 5x20 minute episodes a week, and an ai-generated 30 min super cut of the week's "events" every saturday. What if DragonBallZ, but new episodes drop every morning before your morning commute? And for $30 a month you can do a choose your own adventure where at the end of each episode, the story splits off and you get an alternate history version, b…

Yes, that sounds like one of the ways this can happen. I think children born in the next few years will look back in decades on the cartoons/tv shows/movies they made for themselves.

Re: Ovi: Twin backbone cross-modal fusion for audio-video generation

#106
post #3

mindblowing - but still in the uncanny valley. and I guess it's cute that many of the characters live in a world where AI has caused an apocolypse, but is that really the message they want to lead with?

I watched a few animes which embraced their poor animation and the excellent Dandadan doesnt chase FPS and resolution benchmarks. Wonder if AI video makers will restrain themselves from getting close to the uncanny valley.

Dandadan intro and its lack of FPS and sharp lines: https://www.youtube.com/watch?v=a4na2opArGY

Re: Ovi: Twin backbone cross-modal fusion for audio-video generation

#109
post #3

mindblowing - but still in the uncanny valley. and I guess it's cute that many of the characters live in a world where AI has caused an apocolypse, but is that really the message they want to lead with?

I watched a few animes which embraced their poor animation and the excellent Dandadan doesnt chase FPS and resolution benchmarks. Wonder if AI video makers will restrain themselves from getting close to the uncanny valley. Dandadan intro and its lack of FPS and sharp lines: https://www.youtube.com/watch?v=a4na2opArGY

Dandadan is as far as you can get from “poor animation” but it definitely has a specific mix of lo-fi aesthetic animated exceedingly well, and they’re using Adobe Animate, aka Flash.

Animation doesn’t feel fast if it’s too many FPS or too steady, anyway, ironically and counterintuitively. You can’t do everything on the ones and twos.

To your point about Dandadan’s intro, it’s jam packed with references, which is another kind of skill in and of itself:

https://www.youtube.com/watch?v=5sUaK0xahBU

Chainsaw Man is in that same vein, and is another Science Saru production. I’m looking forward to seeing what they will do with the Ghost in the Shell franchise next year.

https://en.wikipedia.org/wiki/Science_Saru

I get what you mean though regarding Dandadan’s animation style; it has a very hand drawn manga vibe, and the detail is minimal yet finely balanced against the overwhelming amount of noise and visuals. It’s like a slapdash superflat.

https://en.wikipedia.org/wiki/Superflat

On a side note, as an anime fan, MBS is doing great work lately. I liked Witch Watch much more than I expected to, and that’s a much better show than the genres involved would lead one to expect.

https://en.wikipedia.org/wiki/Mainichi_Broadcasting_System

Post reply on HN