Earlier quoted context omitted.
Will larger-parameter versions be released?
We are always figuring out what parameter size makes sense. The decision is always a mix between how good we can make the models from a technical aspect, with how good they need to be to make all of you super excited to use them. And its a bit of a challenge what is an ever changing ecosystem. I'm personally curious is there a certain parameter size you're looking for?
Google releases Gemma 4 open models
141–150 of 507 posts
Re: Google releases Gemma 4 open models
#142Earlier quoted context omitted.
On LM Studio I'm only seeing models/google/gemma-4-26b-a4b Where can I download the full model? I have 128GB Mac Studio
downloading the official ones for my m3 max 128GB via lm studio I can't seem to get them to load. they fail for some unknown reason. have to dig into the logs. any luck for you?
Re: Google releases Gemma 4 open models
#143Earlier quoted context omitted.
On LM Studio I'm only seeing models/google/gemma-4-26b-a4b Where can I download the full model? I have 128GB Mac Studio
downloading the official ones for my m3 max 128GB via lm studio I can't seem to get them to load. they fail for some unknown reason. have to dig into the logs. any luck for you?
Re: Google releases Gemma 4 open models
#144There are so many heavy hitting cracked people like daniel from unsloth and chris lattner coming out of the woodworks for this with their own custom stuff. How does the ecosystem work? Have things converged and standardized enough where it's "easy" (lol, with tooling) to swap out parts such as weights to fit your needs? Do you need to autogen new custom kernels to fix said things? Super cool stuff.
- Lattner tweeted a link to this: https://www.modular.com/blog/day-zero-launch-fastest-perform...
- Unsloth prior post on gemma 3 finetuning: https://unsloth.ai/blog/gemma3
Re: Google releases Gemma 4 open models
#145Gemma 3 was the first model that I have liked enough to use a lot just for daily questions on my 32G gpu.
Re: Google releases Gemma 4 open models
#146Seems like Google and Anthropic (which I consider leaders) would rather keep their secret sauce to themselves – understandable.
Re: Google releases Gemma 4 open models
#147Re: Google releases Gemma 4 open models
#148Downloaded through LM Studio on an M1 Max 32GB, 26B A4B Q4_K_M First message: https://i.postimg.cc/yNZzmGMM/Screenshot-2026-04-03-at-12-44... Not sure if I'm doing something wrong? This more or less reflects my experience with most local models over the last couple years (although admittedly most aren't anywhere near this bad). People keep saying they're useful and yet I can't get them to be consistently useful at al…
Re: Google releases Gemma 4 open models
#149Hi all! I work on the Gemma team, one of many as this one was a bigger effort given it was a mainline release. Happy to answer whatever questions I can
What was the main focus when training this model? Besides the ELO score, it's looking like the models (31B / 26B-A4) are underperforming on some of the typical benchmarks by a wide margin. Do you believe there's an issue with the tests or the results are misleading (such as comparative models benchmaxxing)? Thank you for the release.
You can use this model for about 5 seconds and realize its reasoning is in a league well above any Qwen model, but instead people assume benchmarks that are openly getting used for training are still relevant.
Re: Google releases Gemma 4 open models
#150Google might not have the best coding models (yet) but they seem to have the most intelligent and knowledgeable models of all especially Gemini 3.1 Pro is something. One more thing about Google is that they have everything that others do not: 1. Huge data, audio, video, geospatial 2. Tons of expertise. Attention all you need was born there. 3. Libraries that they wrote. 4. Their own data centers and cloud. 4. Most of…
I recently canceled my Google One subscription because getting accurate answers out of Gemini for chat is basically impossible afaict. Whether I enable thinking makes no difference: Gemini always answers me super quickly, rarely actually looks something up, and lies to me. It has a really bad unchecked hallucination problem because it prioritizes speed over accuracy and (astonishingly, to me) is way more hesitant to…