Live data from Hacker News

Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

github.com

41–50 of 261 posts

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#41

Oh no, someone is profiting off of their work without proper attribution!?!?

"Their work"? First you had the original content creators that did 99.99% of the work. Then you had the US companies bundle it up into a frontier LLM. Then "they" did the "work" of using the US model as a foundation for their own. So in the sense of doing 0.00001% of the actual work that went into their product, sure. I'd say it's more like someone forking a Linux distro, adding a few themes and fonts, and then compl…

Oof this is delete your post level I think. Sorry bud, I been there.

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#43
post #8

Earlier quoted context omitted.

No, typically Brazilians go to Paraguay for their education, most of their technology comes from Paraguay too.

What? Never heard of this

That sounds like nonsense, they don't even speak the same language in Brasil and Paraguay …

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#44

Didn’t the last thread about this have someone from the lab or an enthusiast in Rio saying exactly that? Its a fine tune of Qwen Not a conspiracy

The allegation here is that it's not actually a fine-tune of Qwen, but instead an undisclosed mashup (merge) of someone else's fine-tune of Qwen and the original model. Rio subsequently said that the model was in fact a merge, that they did additional fine-tuning after the merge, and that they accidentally uploaded the base merge instead of the version with additional fine-tuning. But this seems like quite an oversight...

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#45

Oh no, someone is profiting off of their work without proper attribution!?!?

This is a pure scam on tax payer money. But what else would be expected?

Unlike the big companies who do this, which often are merely impure scams on tax payer money a little more downstream.

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#46

One funny thing about incompetence is that they don't have the competence to know that their incompetence is straightforward to verify by a competent person.

I wouldn’t describe what happened here as incompetence. As a “carioca”, I am pleasantly surprised to know that the government’s IT department is involved in AI work — even without the budget to create its own models from scratch.

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#48

Earlier quoted context omitted.

Attribution isn't the relevant part. Lying about your lab's capabilities is.

I do not see anyone lying. The model card says: > Post-trained from Qwen 3.5 397B The model card also says that they use an inference framework based on "SwiReasoning: Switch-Thinking in Latent and Explicit for Pareto-Superior Reasoning LLMs" by Shi et al.: https://arxiv.org/abs/2510.05069 So the sources seem properly attributed. They only claim that what they did to "Qwen 3.5 397B" has improved the LLM, including, a…

Are you talking about the credit that was just updated an hour ago? lol

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#49
post #46

One funny thing about incompetence is that they don't have the competence to know that their incompetence is straightforward to verify by a competent person.

I wouldn’t describe what happened here as incompetence. As a “carioca”, I am pleasantly surprised to know that the government’s IT department is involved in AI work — even without the budget to create its own models from scratch.

This seems kind of insane though, every time I go to Rio I think of the potential of AI/technology to solve some problems and leave it even more paradisiacal... But working on their own model? Wtf? There are a million applications of existing ones there that should be followed up on instead.

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#50

Oh no, someone is profiting off of their work without proper attribution!?!?

Attribution isn't the relevant part. Lying about your lab's capabilities is.

But the whole game is lying and stealing isn't it?
Post reply on HN