Gemma: New Open Models
491–500 of 543 posts
Re: Gemma: New Open Models
#492Earlier quoted context omitted.
This is a TOS, meaning their enforcement option is a lawsuit. In court, if you convincingly argue why it would take an unreasonable amount of effort to update, you win. They can't compel you to unreasonable effort as per their own TOS.
This assumes they even know that the model hasn't been updated. Who is this actually intended for? I'd bet it's for companies hosting the model. In those cases, the definition of reasonable effort is a little closer to "it'll break our stuff if we touch it" rather than "oh silly me, I forgot how to spell r-s-y-n-c".
Re: Gemma: New Open Models
#493Earlier quoted context omitted.
https://twitter.com/yar_vol/status/1760314018575634842
I've worked at Google. It is the organization with highest concentration of engineering talent I've ever been at. Almost to the point that it is ridiculous because you have extremely good engineers working on internal reporting systems for middle managers.
Re: Gemma: New Open Models
#494Earlier quoted context omitted.
https://twitter.com/yar_vol/status/1760314018575634842
Why is this anonymous tweet with no evidence or engagement being posted by multiple users in this thread? Why not just make the same claim directly?
Most programmers are really not that smart nowadays. I’ve seen too many cases of people throwing around claims without a second of deep and critical thought.
Re: Gemma: New Open Models
#495Earlier quoted context omitted.
Any computer program that does not deliver the expected output given a sufficient input is inherently bad.
When Jesus said this: "What father among you, if his son asks for a fish, will instead of a fish give him a serpent?" (Luke 11) He was actually foretelling the future. He saw Gemini.
Re: Gemma: New Open Models
#496Earlier quoted context omitted.
we're at basic knowledge level, if your RAG imply some of it, you can get bad result too. Anyway, would you use a model who makes this nonsense response or one that doesn't? I know which one I will prefer for sure...
If this was better at specific RAG or coding performance I would absolutely, certainly without a doubt use it over a general instruct model in those instances.
Re: Gemma: New Open Models
#497Earlier quoted context omitted.
I understand the theory, I was looking for an example of the same text tokenized with the two different vocabularies.
Do you have an example text in mind? You can use this playground to test it out: https://huggingface.co/spaces/Xenova/the-tokenizer-playgroun...
Doesn't seem like a well made tokenizer at first glance or it's heavily biased towards languages the model can't even generate coherently, lol. If they really wanted it to be SOTA at something they could've at least made it the first open source truly multilingual model, but that's apparently more effort than the lame skin colour oriented virtue signalling Google wants to do.
Re: Gemma: New Open Models
#498I personally can't take any models from google seriously. I was asking it about the Japanese Heian period and it told me such nonsensical information you would have thought it was a joke or parody. Some highlights were "Native American women warriors rode across the grassy plains of Japan, carrying Yumi" and "A diverse group of warriors, including a woman of European descent wielding a katana, stand together in camar…
Re: Gemma: New Open Models
#499Earlier quoted context omitted.
I was wondering if these models would perform in such a way, given this week's X/twitter storm over Gemini generated images. E.g. https://x.com/debarghya_das/status/1759786243519615169?s=20 https://x.com/MiceynComplex/status/1759833997688107301?s=20 https://x.com/AravSrinivas/status/1759826471655452984?s=20
Of all the very very very many things that Google models get wrong, not understanding nationality and skin tone distributions seems to be a very weird one to focus on. Why are there three links to this question? And why are people so upset over it? Very odd, seems like it is mostly driven by political rage.