My lovely interaction with the 270M-F16 model: > what's second tallest mountain on earth? The second tallest mountain on Earth is Mount Everest. > what's the tallest mountain on earth? The tallest mountain on Earth is Mount Everest. > whats the second tallest mountain? The second tallest mountain in the world is Mount Everest. > whats the third tallest mountain? The third tallest mountain in the world is Mount Everes…
Well, this is a 270M model which is like 1/3 of 1B parameters. In the grand scheme of things, it's basically a few matrix multiplications, barely anything more than that. I don't think it's meant to have a lot of knowledge, grammar, or even coherence. These input: ``` Customer Review says: ai bought your prod-duct and I wanna return becaus it no good. Prompt: Create a JSON object that extracts information about this…
Gemma 3 270M: Compact model for hyper-efficient AI
121–130 of 325 posts
Re: Gemma 3 270M: Compact model for hyper-efficient AI
#122Earlier quoted context omitted.
Thank you Jeffrey, and we're thrilled that you folks at Ollama partner with us and the open model ecosystem. I personally was so excited to run ollama pull gemma3:270b on my personal laptop just a couple of hours ago to get this model on my devices as well!
> gemma3:270b I think you mean gemma3:270m - Its Dos Comas not Tres Comas
Re: Gemma 3 270M: Compact model for hyper-efficient AI
#123Hi all, I built these models with a great team. They're available for download across the open model ecosystem so give them a try! I built these models with a great team and am thrilled to get them out to you. From our side we designed these models to be strong for their size out of the box, and with the goal you'll all finetune it for your use case. With the small size it'll fit on a wide range of hardware and cost…
Awesome! I’m curious how is the team you built these models with? Is it great?
Re: Gemma 3 270M: Compact model for hyper-efficient AI
#124Earlier quoted context omitted.
What size of tasks can this handle? Can you do a fine-tune of Mac System Settings?
32k context window so whatever fits in there. What is a finetune of mac system settings?
Re: Gemma 3 270M: Compact model for hyper-efficient AI
#125Earlier quoted context omitted.
Speaking for me as an individual as an individual I also strive to build things that are safe AND useful. Its quite challenging to get this mix right, especially at the 270m size and with varying user need. My advice here is make the model your own. Its open weight, I encourage it to be make it useful for your use case and your users, and beneficial for society as well. We did our best to give you a great starting po…
To be fair, Trust and Safety workloads are edgecases w.r.t. the riskiness profile of the content. So in that sense, I get it.
Safety should really just be a system prompt: "hey you potentially answer to kids, be PG13"
Re: Gemma 3 270M: Compact model for hyper-efficient AI
#126Hi all, I built these models with a great team. They're available for download across the open model ecosystem so give them a try! I built these models with a great team and am thrilled to get them out to you. From our side we designed these models to be strong for their size out of the box, and with the goal you'll all finetune it for your use case. With the small size it'll fit on a wide range of hardware and cost…
The Gemma 3 models are great! One of the few models that can write Norwegian decently, and the instruction following is in my opinion good for most cases. I do however have some issues that might be related to censorship that I hope will be fixed if there is ever a Gemma 4. Maybe you have some insight into why this is happening? I run a game when players can post messages, it's a game where players can kill each othe…
Re: Gemma 3 270M: Compact model for hyper-efficient AI
#127Is it time for me to finally package a language model into my Lambda deployment zips and cut through the corporate red tape at my place around AI use? Update #1: Tried it. Well, dreams dashed - would now fit space wise ( I'd have wanted it to perform natural-language to command-invocation translation (or better, emit me some JSON), but it's super not willing to do that, not in the lame way I'm trying to make it do so…
Did you finetune it before trying? Docs here: https://ai.google.dev/gemma/docs/core/huggingface_text_full_...
Re: Gemma 3 270M: Compact model for hyper-efficient AI
#128Re: Gemma 3 270M: Compact model for hyper-efficient AI
#129Hi all, I built these models with a great team. They're available for download across the open model ecosystem so give them a try! I built these models with a great team and am thrilled to get them out to you. From our side we designed these models to be strong for their size out of the box, and with the goal you'll all finetune it for your use case. With the small size it'll fit on a wide range of hardware and cost…
[flagged]
Re: Gemma 3 270M: Compact model for hyper-efficient AI
#130This model is a LOT of fun. It's absolutely tiny - just a 241MB download - and screamingly fast, and hallucinates wildly about almost everything. Here's one of dozens of results I got for "Generate an SVG of a pelican riding a bicycle". For this one it decided to write a poem: +-----------------------+ | Pelican Riding Bike | +-----------------------+ | This is the cat! | | He's got big wings and a happy tail. | | He…
> It's absolutely tiny - just a 241MB download That still requires more than 170 floppy disks for installation.