Oh no, someone is profiting off of their work without proper attribution!?!?
"Their work"? First you had the original content creators that did 99.99% of the work. Then you had the US companies bundle it up into a frontier LLM. Then "they" did the "work" of using the US model as a foundation for their own. So in the sense of doing 0.00001% of the actual work that went into their product, sure. I'd say it's more like someone forking a Linux distro, adding a few themes and fonts, and then compl…
Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
41–50 of 261 posts
Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
#42Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
#43Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
#44Didn’t the last thread about this have someone from the lab or an enthusiast in Rio saying exactly that? Its a fine tune of Qwen Not a conspiracy
Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
#45Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
#46One funny thing about incompetence is that they don't have the competence to know that their incompetence is straightforward to verify by a competent person.
Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
#47-- Bill Gates
Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
#48Earlier quoted context omitted.
Attribution isn't the relevant part. Lying about your lab's capabilities is.
I do not see anyone lying. The model card says: > Post-trained from Qwen 3.5 397B The model card also says that they use an inference framework based on "SwiReasoning: Switch-Thinking in Latent and Explicit for Pareto-Superior Reasoning LLMs" by Shi et al.: https://arxiv.org/abs/2510.05069 So the sources seem properly attributed. They only claim that what they did to "Qwen 3.5 397B" has improved the LLM, including, a…
Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
#49One funny thing about incompetence is that they don't have the competence to know that their incompetence is straightforward to verify by a competent person.
I wouldn’t describe what happened here as incompetence. As a “carioca”, I am pleasantly surprised to know that the government’s IT department is involved in AI work — even without the budget to create its own models from scratch.