Live data from Hacker News

AniSora: Open-source anime video generation model

komiko.app

171–180 of 232 posts

Re: AniSora: Open-source anime video generation model

#171
post #92

Earlier quoted context omitted.

> As my mom retired from being a translator, she went from typewriter to machine-assisted translation with centralised corpus-databases. All the while the available work became less and less, and the wages became lower and lower. She was lucky to be able to retire when she did, as the job of a translator is definitely going to become extinct. You can already get higher quality translations from machine learning model…

While LLMs are pretty good, and likely to improve, my experience is OpenAI's offerings *absolutely* make stuff up after a few thousand words or so, and they're one of the better ones. It also varies by language. Every time I give an example here of machine translated English-to-Chinese, it's so bad that the responses are all people who can read Chinese being confused because it's gibberish. And as for politics, as Gr…

> While LLMs are pretty good, and likely to improve, my experience is OpenAI's offerings absolutely make stuff up after a few thousand words or so, and they're one of the better ones.

That's not how you get good translations from off-the-shelf LLMs! If you give a model the whole book and expect it to translate it in one-shot then it will eventually hallucinate and give you bad results.

What you want is to give it a small chunk of text to translate, plus previously translated context so that it can keep the continuity.

And for the best quality translations what you want is to use a dedicated model that's specifically trained for your language pairs.

> And as for politics, as Grok has just been demonstrating, they're quite capable of whatever bias they've been trained to have or told to express.

In an open ended questions - sure. But that doesn't apply to translations where you're not asking the model to come up with something entirely by itself, but only getting it to accurately translate what you wrote into another language.

I can give you an example. Let's say we want to translate the following sentence:

"いつも言われるから、露出度抑えたんだ。"

Let's ask a general purpose LLMs to translate it without any context (you could get a better translation if you'd give it context and more instructions):

ChatGPT (1): "Since people always comment on it, I toned down how revealing it is."

ChatGPT (2): "People always say something, so I made it less revealing."

Qwen3-235B-A22B: "I always get told, so I toned down how revealing my outfit is."

gemma-3-27b-it (1): "Because I always get told, I toned down how much skin I show."

gemma-3-27b-it (2): "Since I'm always getting comments about it, I decided to dress more conservatively."

gemma-3-27b-it (3): "I've been told so often, I decided to be more modest."

Grok: "I was always told, so I toned down the exposure."

And how humans would translate it:

Competent human translator (I can confirm this is an accurate translation, but perhaps a little too literal): "Everyone was always saying something to me, so I tried toning down the exposure."

Activist human translator: "Oh those pesky patriarchal societal demands were getting on my nerves, so I changed clothes."

(Source: https://www.youtube.com/watch?v=dqaAgAyBFQY)

It should be fairly obvious which one is the biased one, and I don't think it's the Grok one (which is a little funny, because it's actually the most literal translation of them all).

Re: AniSora: Open-source anime video generation model

#174

I welcome this. I know there is a huge market for those excited for infinite anime music videos and all things anime. This is great for an abundance of content and everyone will become anime artists now. Japan is truly is embracing AI and there will be new jobs for everyone thanks to the boom AI is creating as well as Jevons paradox which will create huge demand. Even better if this open source.

I don't know, I used to like some anime and mangas when I was 14 in the mid 90's. Nowadays it seems everyone is interested by "anime style" of content but all I see is terrible in term of quality. It seems quantity increased so much in the last 30 years it only made quality stuff more invisible and we are inundated with animelike trash.

The percentage of anime I like is low and has always been low. I find a new anime I like comes along about every three years (I have to dig for it though.) In general, I care about the writing and story more than the visuals. So with a great increase in the amount of anime a single writer can create, shouldn't this allow for more well-written sloppy-visuals anime to exist? I'm excited to see.

Re: AniSora: Open-source anime video generation model

#175
post #70

Why? Who needs this? Who wants this? I still don't get why you would produce art with generation models instead of letting human artists do their thing. It's only funny as long as it's bad, but once it becomes better it's just creepy and most of all totally pointless.

Why is it pointless? People want anime. This technology allows more anime to exist. It's like you're saying "why do we need cast-iron moulds? Just let artisans do their craft."

Re: AniSora: Open-source anime video generation model

#178

Earlier quoted context omitted.

The way I think of art has two main components: the aesthetics and the higher level impressions invoked through those aesthetics. For me, art is specifically about the experimentation-with and the then-intentional use of aesthetics, to evoke a specific set of impressions within its audience. A form of communication, a transfer of experiences, frames of mind, and ideas. The more effectively and intelligently one can d…

It's human intent. AI is technically another tool, and it can be used poorly (what people refer to "AI slop", using default settings, some LoRA and calling it a day) and it can be used properly (forcing compositions, editing, fixing errors...) to convey an idea or emotion or tell a story. Critical eye does the rest. After all, the machine doesn't do anything on its own, it needs a driver. The quality of the output is…

Sure, but intent is a very fickle thing.

Consider zero and single click deployments in IT operations. With single click deployments, you need to have everything automated, but the go sign is still given by a human. With zero click, you'll have a deployment policy instead - the human decision is now out of the critical path completely, and only plays part during the authoring and later editing of said policy. And you can also then generate those policies, and so on.

Same can be applied to AI. You can have canned prompts that you keep refining to encode your intent and preferences, then you just use them with a single click. But you can also build a harness that generates prompts and continuously monitors trends and the world as a whole for any kind of arbitrary criteria (potentially of its own random or even shifting choice), and then follows that: a reward policy. And then like with regular IT, you can keep layering onto that.

Because of this, I don't think that intent is the point of differentiation necessarily, but the experience and shared understanding of human intent. That people have varying individual, arbitrary preferences, and are going through life of arbitrary and endless differences, and then source from those to then create. Indeed, this is never going to be replicated, exactly because of what I said: this is humans being human, and that giving them a unique, inalienable position by definition.

It's like if instead of planes we called aircraft "mechanical birds" and dunked on them for not flying by flapping their wings, despite their more than a century long history and claims of "flying". But just like I think planes do actually fly, I do also think that these models produce art. [0]

[0] https://youtu.be/ipRvjS7q1DI

Re: AniSora: Open-source anime video generation model

#180
post #71

Earlier quoted context omitted.

You can't compare translation to creating new works of art. Sorry mom, but that's apples and oranges. A dangerously false comparison.

If you speak more than one language(esp something like Chinese or Japanese) you understand how subjective some choices are. It certainly takes creative decision making.

I speak Japanese natively and hell, I'm just going to say, there is no such thing as translation, there is just foreign language ghostwriting.

I'm not even sure if bilingualism is real or if it's just an alternate expression for relatively benign forced split personality. Could very well be.

Post reply on HN