Live data from Hacker News

Gemma: New Open Models

blog.google

1–10 of 543 posts

Re: Gemma: New Open Models

#5

Hello on behalf of the Gemma team! We are really excited to answer any questions you may have about our models. Opinions are our own and not of Google DeepMind.

Congrats on the launch and thanks for the contribution! This looks like it's on-par or better compared to mistral 7B 0.1 or is that 0.2?

Are there plans for MoE or 70B models?

Re: Gemma: New Open Models

#7

Hello on behalf of the Gemma team! We are really excited to answer any questions you may have about our models. Opinions are our own and not of Google DeepMind.

How are these performing so well compared to Llama 2, are there any documents on the architecture and differences, is it MoE?

Also note some of the links on the blog post don't work, e.g debugging tool.

Re: Gemma: New Open Models

#8

Hello on behalf of the Gemma team! We are really excited to answer any questions you may have about our models. Opinions are our own and not of Google DeepMind.

It seems you have exposed the internal debugging tool link in the blog post. You may want to do something about it.
Post reply on HN