Live data from Hacker News

Gemma 4 on iPhone

apps.apple.com

121–130 of 267 posts

Re: Gemma 4 on iPhone

#121

Earlier quoted context omitted.

Did you really watch “Her” and think this is a future that should happen?? Seriously????

What does what they said have anything to do with Her? Local LLMs are better than big corporations owning your data and offering LLMs for a huge cost.

I get the local ai thing. I agree it’s probably a good direction. The bit that has to do with the movie “her” is the bit at the end where they are excited about “her”-like companions on our phones.

Re: Gemma 4 on iPhone

#122
post #42

Earlier quoted context omitted.

Did you really watch “Her” and think this is a future that should happen?? Seriously????

I don’t think OP’s point has anything to do with AI companions. The big benefit of moving compute to edge devices is to distribute the inference load on the grid. Powering and cooling phones is a lot easier than powering and cooling a datacenter

Local ai is probably a good direction, i agree. But there was a part of their point that had to do with ai companions: the bit where they say we are closer to “her”-like ai companions. That was the bit i was responding to.

Re: Gemma 4 on iPhone

#123
post #57

Earlier quoted context omitted.

> or in the cloud but way more expensive then it is today. Why? It's widely understood that the big players are making profit on inference. The only reason they still have losses is because training is so expensive, but you need to do that no matter whether the models are running in the cloud or on your device. If you think about it, it's always going to be cheaper and more energy-efficient to have dedicated cloud ha…

> It's widely understood that the big players are making profit on inference. This is most definitely not widely understood. We still don't know yet. There's tons of discussions about people disagreeing on whether it really is profitable. Unless you have proof, don't say "this is widely understood".

The reality is we can’t trust accounting earnings anyway.

We need to see the cash flows.

Re: Gemma 4 on iPhone

#128

Earlier quoted context omitted.

> or in the cloud but way more expensive then it is today. Why? It's widely understood that the big players are making profit on inference. The only reason they still have losses is because training is so expensive, but you need to do that no matter whether the models are running in the cloud or on your device. If you think about it, it's always going to be cheaper and more energy-efficient to have dedicated cloud ha…

> It's widely understood that the big players are making profit on inference. If you add in the cost of training, it’s not profitable. Not including the cost of training is a bit like saying the only cost of a cup of coffee is the paper cup it’s in. The only way OpenAI gets to charge for inference is by selling a product people can’t get elsewhere for much cheaper, which means billions in R&D costs. But because of co…

At least Anthropic claims that they are profitable on a per model basis. But since both revenue and training costs are growing exponentially, and they need to pay for model N training today, and only get revenue for model N-1 today, the offset makes it look worse than it is.

Obviously that doesn’t help them turn a profit, until they can stop growing training costs exponentially.

So it’s really a race to see whether growth in revenue or training costs decelerates first.

Re: Gemma 4 on iPhone

#129

Nice! Tried on iPhone 16 pro with 30 TPS from Gemma-4-E2B-it model. Although the phone got considerably hot while inferencing. It’s quite an impressive performance and cannot wait to try it myself in one of my personal apps.

It's at least somewhat limited in non-English content. It knows how to make lentil soup, so I was happy that I never need to look up recipe sites with awful UX and ads, but then it couldn't find a recipe for "Kalter Hund"/"Kalte Schnauze". So sad ;)

Still, absolutely fabulous. What a time to be alive!

Post reply on HN