Live data from Hacker News

Gemma: New Open Models

blog.google

491–500 of 543 posts

Re: Gemma: New Open Models

#492

Earlier quoted context omitted.

This is a TOS, meaning their enforcement option is a lawsuit. In court, if you convincingly argue why it would take an unreasonable amount of effort to update, you win. They can't compel you to unreasonable effort as per their own TOS.

This assumes they even know that the model hasn't been updated. Who is this actually intended for? I'd bet it's for companies hosting the model. In those cases, the definition of reasonable effort is a little closer to "it'll break our stuff if we touch it" rather than "oh silly me, I forgot how to spell r-s-y-n-c".

Hosting companies can probably just claim they're covered under Section 230, and Google has to go bother the individual users, not them.

Re: Gemma: New Open Models

#493
post #363
post #193

Earlier quoted context omitted.

https://twitter.com/yar_vol/status/1760314018575634842

I've worked at Google. It is the organization with highest concentration of engineering talent I've ever been at. Almost to the point that it is ridiculous because you have extremely good engineers working on internal reporting systems for middle managers.

GDM works on internal systems? It’s the first time I hear this.

Re: Gemma: New Open Models

#494
post #193

Earlier quoted context omitted.

https://twitter.com/yar_vol/status/1760314018575634842

Why is this anonymous tweet with no evidence or engagement being posted by multiple users in this thread? Why not just make the same claim directly?

Programming popularized -> more people -> more cases of knee jerk reaction encountered.

Most programmers are really not that smart nowadays. I’ve seen too many cases of people throwing around claims without a second of deep and critical thought.

Re: Gemma: New Open Models

#495

Earlier quoted context omitted.

Any computer program that does not deliver the expected output given a sufficient input is inherently bad.

When Jesus said this: "What father among you, if his son asks for a fish, will instead of a fish give him a serpent?" (Luke 11) He was actually foretelling the future. He saw Gemini.

Hahaha. The man had a lot of wisdom, after all.

Re: Gemma: New Open Models

#496
post #336

Earlier quoted context omitted.

we're at basic knowledge level, if your RAG imply some of it, you can get bad result too. Anyway, would you use a model who makes this nonsense response or one that doesn't? I know which one I will prefer for sure...

If this was better at specific RAG or coding performance I would absolutely, certainly without a doubt use it over a general instruct model in those instances.

People getting so used to being manipulated and lied to that they don't even bother anymore is a huge part of the problem. But sure, do what suits you the best.

Re: Gemma: New Open Models

#497

Earlier quoted context omitted.

I understand the theory, I was looking for an example of the same text tokenized with the two different vocabularies.

Do you have an example text in mind? You can use this playground to test it out: https://huggingface.co/spaces/Xenova/the-tokenizer-playgroun...

Interesting, it's actually worse than GPT-4s 100k tokenizer by quite a bit despite being over twice the size and only marginally better than LLama's 30k. At least for some random articles in English latin script that I tried anyway, but Llama and Gemma are English-only models so no point in testing anything else.

Doesn't seem like a well made tokenizer at first glance or it's heavily biased towards languages the model can't even generate coherently, lol. If they really wanted it to be SOTA at something they could've at least made it the first open source truly multilingual model, but that's apparently more effort than the lame skin colour oriented virtue signalling Google wants to do.

Re: Gemma: New Open Models

#498
post #257

I personally can't take any models from google seriously. I was asking it about the Japanese Heian period and it told me such nonsensical information you would have thought it was a joke or parody. Some highlights were "Native American women warriors rode across the grassy plains of Japan, carrying Yumi" and "A diverse group of warriors, including a woman of European descent wielding a katana, stand together in camar…

How are you running the model? I believe it's a bug from a rushed instruct fine-tuning or in the chat template. The base model can't possibly be this bad. https://github.com/ollama/ollama/issues/2650

Re: Gemma: New Open Models

#499
post #279

Earlier quoted context omitted.

I was wondering if these models would perform in such a way, given this week's X/twitter storm over Gemini generated images. E.g. https://x.com/debarghya_das/status/1759786243519615169?s=20 https://x.com/MiceynComplex/status/1759833997688107301?s=20 https://x.com/AravSrinivas/status/1759826471655452984?s=20

Of all the very very very many things that Google models get wrong, not understanding nationality and skin tone distributions seems to be a very weird one to focus on. Why are there three links to this question? And why are people so upset over it? Very odd, seems like it is mostly driven by political rage.

Maybe some people care about truth?

Re: Gemma: New Open Models

#500
Hopefully, they re-release this under an open license. Making everyone go through the exercise to authenticate and agree to terms hasn't worked for any model to-date. It just limits it's reach. We saw the same thing with Phi-2.
Post reply on HN