Live data from Hacker News

Gemma 3 Technical Report [pdf]

storage.googleapis.com

121–130 of 260 posts

Re: Gemma 3 Technical Report [pdf]

#121
post #115

Earlier quoted context omitted.

All of this is true but then it's as easy as releasing censored and uncensored versions of the model. Then it's up to users (or parents, in the case of children) to choose the adequate version for each purpose. Just like there are child-friendly movies and adult-only movies, and no one beyond fringe puritan crusaders would say that the latter should outright not exist.

>censored and uncensored Well here you still have the same problem, since they're not gonna release an actually uncensored version, that tells you how to do awful things (or indeed, that tells you to do them). So then you'd have censored and less censored, and it would still be a matter of where to draw those lines.

True, "uncensored" is not the best term for what I meant (as I'm aware that fully uncensored is not a realistic thing to ask from companies).

What I mean is a model for all audiences and an adult model, and the line would be drawn at the law of the country producing it (if it's something that would be legal to publish for a human author at a website, then it should be allowed as an LLM response). So erotica would be fine, while instructions for making a bomb wouldn't.

Re: Gemma 3 Technical Report [pdf]

#122
Thanks for these cool models!

One suggestion (or just rant): Less censorship for local models, PLEASE.

One question: 100+ elo gains from gemma 2 to gemma 3 on Chatbot arena is really something, any estimates on how this is achieved?

Re: Gemma 3 Technical Report [pdf]

#123
If it helps anyone, I wrote a detailed analysis here: https://x.com/danielhanchen/status/1899735308180267176

TLDR:

1. 1B text only, 4, 12, 27B Vision + text. 14T tokens

2. 128K context length further trained from 32K. 1B is 32K.

3. Removed attn softcapping. Replaced with QK norm

4. 5 sliding + 1 global attn

5. 1024 sliding window attention

6. RL - BOND, WARM, WARP

Re: Gemma 3 Technical Report [pdf]

#124

Earlier quoted context omitted.

Hey, I'm Ravin from the Gemma team. It's on ollama! Try `ollama run gemma3` to get it pulled locally

My point was multi-images and pan-and-scan. We haven't implemented those yet in Ollama, but soon!

Good, FYI the number one usage is vision RAGs (RAGs that deal with documents as images instead of text).

Re: Gemma 3 Technical Report [pdf]

#125

Gemma 3 is out! Multimodal (image + text), 128K context, supports 140+ languages, and comes in 1B, 4B, 12B, and 27B sizes with open weights & commercial use. Gemma 3 model overview: https://ai.google.dev/gemma/docs/core Huggingface collection: https://huggingface.co/collections/google/gemma-3-release-67... ollama: https://ollama.com/library/gemma3

I'm still a huge fan of gemma-22b. Looking forward to this one!

Re: Gemma 3 Technical Report [pdf]

#126

> They are designed to help prevent our models from generating harmful content, i.e., > [...] > Sexually explicit content Dear tech companies. Sexually explicit content is not harmful. Why are you all run by puritans? I don't even want to make edgy porn, I just want to be treated like an adult.

Whenever they say things like "harmful" or "unsafe" there is an implied "for our brand" that follows.

Re: Gemma 3 Technical Report [pdf]

#127

> They are designed to help prevent our models from generating harmful content, i.e., > [...] > Sexually explicit content Dear tech companies. Sexually explicit content is not harmful. Why are you all run by puritans? I don't even want to make edgy porn, I just want to be treated like an adult.

It's harmful in that there exists a significant and vocal subset of users who does not wish to see that content or does not wish their children to do so. It's easier to teach your model never to produce that kind of content than to teach it to perfectly distinguish whether this user should see that content or not. TV channels are barred from broadcasting this kind of content for similar reasons. Sure, there are alway…

Yes, it would be absolutely shameful if there was pornography on the internet, easily available to anyone, even children. Society would crumble!

Re: Gemma 3 Technical Report [pdf]

#128
post #64

Greetings from the Gemma team! We just got Gemma 3 out of the oven and are super excited to show it to you! Please drop any questions here and we'll answer ASAP. (Opinions our own and not of Google DeepMind.) PS we are hiring: https://boards.greenhouse.io/deepmind/jobs/6590957

will there ever be a Gemma 3 Thinking? how copyable is the Flash Thinking approach to the Gemma series?

That's a very interesting area, but nothing we can announce today.

Re: Gemma 3 Technical Report [pdf]

#129

Earlier quoted context omitted.

It's harmful in that there exists a significant and vocal subset of users who does not wish to see that content or does not wish their children to do so. It's easier to teach your model never to produce that kind of content than to teach it to perfectly distinguish whether this user should see that content or not. TV channels are barred from broadcasting this kind of content for similar reasons. Sure, there are alway…

Yes, it would be absolutely shameful if there was pornography on the internet, easily available to anyone, even children . Society would crumble!

Porn sites are blocked in many jurisdictions, so I would not use that argument.

Re: Gemma 3 Technical Report [pdf]

#130
post #129

Earlier quoted context omitted.

Yes, it would be absolutely shameful if there was pornography on the internet, easily available to anyone, even children . Society would crumble!

Porn sites are blocked in many jurisdictions, so I would not use that argument.

No, there's no movement to shut down pornography on the internet. There's a movement to shut down specific websites and make a lot of noise about it but continue consuming pornography behind closed doors.

People like pornography. They'll as soon ban alcohol again (which worked so well last time)

Post reply on HN