Live data from Hacker News

Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

github.com

11–20 of 261 posts

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#14

Oh no, someone is profiting off of their work without proper attribution!?!?

Attribution isn't the relevant part. Lying about your lab's capabilities is.

That's also something all the AI companies have been doing.

Re: Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model

#20

This is fascinating that it worked though. Can we just merge all the open weight models and get something better?

that kinda worked in llama 1/2 era, not between different models but between finetunes of the same model. the briefly legendary Mythomax was IIRC a merge of 5+ tunes, some of which were merges themselves.
Post reply on HN