I have tried a lot of local models. I have 656GB of them on my computer so I have experience with a diverse array of LLMs. Gemma has been nothing to write home about and has been disappointing every single time I have used it. Models that are worth writing home about are; EXAONE-3.5-7.8B-Instruct - It was excellent at taking podcast transcriptions and generating show notes and summaries. Rocinante-12B-v2i - Fun for s…
Gemma3 – The current strongest model that fits on a single GPU
21–30 of 148 posts
Re: Gemma3 – The current strongest model that fits on a single GPU
#22Re: Gemma3 – The current strongest model that fits on a single GPU
#23Not sure if anyone else experiences this, but ollama downloads starts off strong but the last few MBs take forever. Finally just finished downloading (gemma3:27b). Requires the latest version of Ollama to use, but now working, getting about 21 tok/s on my local 2x A4000. From my few test prompts looks like a quality model, going to run more tests to compare against mistral-small:24b to see if it's going to become my…
Re: Gemma3 – The current strongest model that fits on a single GPU
#24After reading the technical report do the effort of downloading the model and run it against a few prompts. In 5 minutes you understand how broken LLM benchmarking is.
Re: Gemma3 – The current strongest model that fits on a single GPU
#25These bar charts are getting more disingenuous every day. This one makes it seem like Gemma3 ranks as nr. 2 on the arena just behind the full DeepSeek R1. But they just cut out everything that ranks higher. In reality, R1 currently ranks as nr. 6 in terms of Elo. It's still impressive for such a small model to compete with much bigger models, but at this point you can't trust any publication by anyone who has any ski…
Re: Gemma3 – The current strongest model that fits on a single GPU
#26After reading the technical report do the effort of downloading the model and run it against a few prompts. In 5 minutes you understand how broken LLM benchmarking is.
can you expand a bit?
Re: Gemma3 – The current strongest model that fits on a single GPU
#27Re: Gemma3 – The current strongest model that fits on a single GPU
#28I have tried a lot of local models. I have 656GB of them on my computer so I have experience with a diverse array of LLMs. Gemma has been nothing to write home about and has been disappointing every single time I have used it. Models that are worth writing home about are; EXAONE-3.5-7.8B-Instruct - It was excellent at taking podcast transcriptions and generating show notes and summaries. Rocinante-12B-v2i - Fun for s…
IME Qwen2.5-3B-Instruct (or even 1.5B) have been quite remarkable, but I haven't done that heavy testing.
Re: Gemma3 – The current strongest model that fits on a single GPU
#29Is "OpenAI" the only AI company that hasn't released any model weights?