Gemma, Gemini pro, Gemini advanced, Gemini ultra
To a layperson it is not obvious which one is better than the other
181–190 of 543 posts
Gemma, Gemini pro, Gemini advanced, Gemini ultra
To a layperson it is not obvious which one is better than the other
Available on Ollama?
Hello on behalf of the Gemma team! We are really excited to answer any questions you may have about our models. Opinions are our own and not of Google DeepMind.
Will there be "extended context" releases like 01.ai did for Yi? Also, is the model GQA?
The utter bullshit of these licenses has got to stop. Do not, under any circumstances, consider using these commercially. "Google reserves the right to restrict (remotely or otherwise) usage of any of the Gemma Services that Google reasonably believes are in violation of this Agreement." This is a kill switch that Google maintains in perpetuity over any system you build relying on these models. Our legal review of th…
Could you share what models you consider to be OK for commercialization?
Hopefully not totally gimped like Gemini. Are they releasing an uncensored version?
Earlier quoted context omitted.
Honestly, this is more of a PR stunt to advertise the Google Dev ecosystem than a contribution to open-source. I'm not complaining, just calling it what it is. Barely an improvement over the 5-month-old Mistral model, with the same context length of 8k. And this is a release after their announcement of Gemini Pro 1.5, which had an exponential increase in context length.
mistral 7b v0.2 supports 32k
I think so many people (including me) effectively ignored Mistral 0.1's sliding window that few realized 0.2 instruct is native 32K.
Earlier quoted context omitted.
Strong disagree - a Mistral fine tune of llama 70b was the top performing llama fine tune. They have lots of data the community simply does not.
Miqu was (allegedly) an internal continued pretrain Mistral did as a test, that was leaked as a GGUF. Maybe its just semantics, it is technically a finetune... But to me theres a big difference between expensive "continuation training" (like Solar 10.7B or Mistral 70B) and a much less intense finetuning. The former is almost like releasing a whole new base model. It would be awesome if Mistral did that with their dat…
I wonder if people will get confused with the naming Gemma, Gemini pro, Gemini advanced, Gemini ultra To a layperson it is not obvious which one is better than the other
Hello on behalf of the Gemma team! We are really excited to answer any questions you may have about our models. Opinions are our own and not of Google DeepMind.