Live data from Hacker News

Gemma 3n preview: Mobile-first AI

developers.googleblog.com

121–130 of 179 posts

Re: Gemma 3n preview: Mobile-first AI

#121
post #92

tried out google/gemma-3n-E4B-it-litert-preview on galaxy s25 ultra loads pretty fast. starts to reply near-instant (text chat mode). doesn't answer questions like "when is your cutoff date" apparently answers "may 15 2024" as today date so probably explains why it answered joe biden as answer to who is US president

I always get a little confused when people fact-check bare foundation models. I don't consider them as fact-bearing, but only fact-preserving when grounded in context.

Am I missing something?

Re: Gemma 3n preview: Mobile-first AI

#122

Earlier quoted context omitted.

I did the same thing on my Pixel Fold. Tried two different images with two different prompts: "What can you see?" and "Describe this image" First image ('Describe', photo of my desk) - 15.6 seconds to first token - 2.6 tokens/second - Total 180 seconds Second image ('What can you see?', photo of a bowl of pasta) - 10.3 seconds to first token - 3.1 tokens/second - Total 26 seconds The Edge Gallery app defaults to CPU…

Pixel 4a release date = August 2020 Pixel Fold was in the Pixel 8 generation but uses the Tensor G2 from the 7s. Pixel 7 release date = October 2022 That's a 26 month difference, yet a full order of magnitude difference in token generation rate on the CPU. Who said Moore's Law is dead? ;)

8 has G3 chip

Re: Gemma 3n preview: Mobile-first AI

#124
post #13

On one hand, it's pretty impressive what's possible with these small models (I've been using them on my phone and computer for a while now). On the other hand, I'm really not looking forward to app sizes ballooning even more – there's no reasonable way to share them across apps at least on iOS, and I can absolutely imagine random corporate apps to start including LLMs, just because it's possible.

Windows is adding an OS-level LLM (Copilot), Chrome is adding a browser-level LLM (Gemini), it seems like Android is gearing up to add an OS-level LLM (Gemmax), and there are rumors the next game consoles might also have an OS-level LLM. It feels inevitable that we'll eventually get some local endpoints that let applications take advantage of on-device generations without bundling their own LLM -- hopefully.

Re: Gemma 3n preview: Mobile-first AI

#125

You can try it on Android right now: Download the Edge Gallery apk from github: https://github.com/google-ai-edge/gallery/releases/tag/1.0.0 Download one of the .task files from huggingface: https://huggingface.co/collections/google/gemma-3n-preview-6... Import the .task file in Edge Gallery with the + bottom right. You can take pictures right from the app. The model is indeed pretty fast.

Suggest giving it no networking permissions (if indeed this is about on-device AI).

Re: Gemma 3n preview: Mobile-first AI

#127
post #13

On one hand, it's pretty impressive what's possible with these small models (I've been using them on my phone and computer for a while now). On the other hand, I'm really not looking forward to app sizes ballooning even more – there's no reasonable way to share them across apps at least on iOS, and I can absolutely imagine random corporate apps to start including LLMs, just because it's possible.

Windows is adding an OS-level LLM (Copilot), Chrome is adding a browser-level LLM (Gemini), it seems like Android is gearing up to add an OS-level LLM (Gemmax), and there are rumors the next game consoles might also have an OS-level LLM. It feels inevitable that we'll eventually get some local endpoints that let applications take advantage of on-device generations without bundling their own LLM -- hopefully.

> It feels inevitable that we'll eventually get some local endpoints that let applications take advantage of on-device generations without bundling their own LLM -- hopefully.

Given the most "modern" and "hip" way of shipping desktop applications seems to be for everyone and their mother to include a browser runtime together with their 1MB UI, don't get your hopes up.

Re: Gemma 3n preview: Mobile-first AI

#128
post #86

Earlier quoted context omitted.

It seems straightforward to me: Apps take up storage, and the only way to get more of that is to pay Apple's markups, as iOS devices don't support upgradable storage. On top of the already hefty markup, they don't even take storage capacity into consideration for trade-ins.

I am not aware of any phone allowing storage upgrades.

Probably most phones before the iPhone (and many afterwards) had SD cards support, you've really never came across that? I remember my Sony Ericsson around 2004 or something had support for it even.

Re: Gemma 3n preview: Mobile-first AI

#129

Earlier quoted context omitted.

I can't speak for anyone else, but these models only seem about as smart as google search, with enormous variability. I can't say I've ever had an interaction with a chatbot that's anything redolent of interaction with intelligence. Now would I take AI as a trivia partner? Absolutely. But that's not really the same as what I look for in "smart" humans.

The image description capabilities are pretty insane, crazy to think it's all happening on my phone. I can only imagine how interesting this is accessibility wise, e.g. for vision impaired people. I believe there are many more possible applications for these on a smartphone than just chatting with them.

[deleted]

Re: Gemma 3n preview: Mobile-first AI

#130
Absolute shit. Comparing it to Sonnet 3.7 is an insult.

# Is Eiffel Tower or a soccer ball bigger ?

> A soccer ball is bigger than the Eiffel Tower! Here's a breakdown:

> Eiffel Tower: Approximately 330 meters (1,083 feet) tall.

> Soccer Ball: A standard soccer ball has a circumference of about 68-70 cm (27-28 inches).

> While the Eiffel Tower is very tall, its base is relatively small compared to its height. A soccer ball, though much smaller in height, has a significant diameter, making it physically larger in terms of volume.

Post reply on HN