Its a fine tune of Qwen
Not a conspiracy
21–30 of 261 posts
Its a fine tune of Qwen
Not a conspiracy
Oh no, someone is profiting off of their work without proper attribution!?!?
Attribution isn't the relevant part. Lying about your lab's capabilities is.
The model card says:
> Post-trained from Qwen 3.5 397B
The model card also says that they use an inference framework based on "SwiReasoning: Switch-Thinking in Latent and Explicit for Pareto-Superior Reasoning LLMs" by Shi et al.:
https://arxiv.org/abs/2510.05069
So the sources seem properly attributed.
They only claim that what they did to "Qwen 3.5 397B" has improved the LLM, including, as expected, with "strong performance in Portuguese".
Oh no, someone is profiting off of their work without proper attribution!?!?
I'd say it's more like someone forking a Linux distro, adding a few themes and fonts, and then complaining when someone else forks their distro and adds another theme.
Oh no, someone is profiting off of their work without proper attribution!?!?
Oh no, someone is profiting off of their work without proper attribution!?!?
This is fascinating that it worked though. Can we just merge all the open weight models and get something better?
also only work on matching architectures (i.e. finetunes/loras of the same model)
Oh no, someone is profiting off of their work without proper attribution!?!?
Oh no, someone is profiting off of their work without proper attribution!?!?
The municipality of Rio de Janeiro (via its IT company IplanRIO) released Rio-3.5-Open-397B, presented as a homegrown Qwen3.5 fine-tune that beats comparable open models on benchmarks. The linked issue argues it's actually a weighted merge of ~60% Nex-N2 Pro + ~40% Qwen3.5-397B-A17B - Nex-N2 having been released about a week earlier.