Earlier quoted context omitted.
Did you really watch “Her” and think this is a future that should happen?? Seriously????
What does what they said have anything to do with Her? Local LLMs are better than big corporations owning your data and offering LLMs for a huge cost.
Gemma 4 on iPhone
121–130 of 267 posts
Re: Gemma 4 on iPhone
#122Earlier quoted context omitted.
Did you really watch “Her” and think this is a future that should happen?? Seriously????
I don’t think OP’s point has anything to do with AI companions. The big benefit of moving compute to edge devices is to distribute the inference load on the grid. Powering and cooling phones is a lot easier than powering and cooling a datacenter
Re: Gemma 4 on iPhone
#123Earlier quoted context omitted.
> or in the cloud but way more expensive then it is today. Why? It's widely understood that the big players are making profit on inference. The only reason they still have losses is because training is so expensive, but you need to do that no matter whether the models are running in the cloud or on your device. If you think about it, it's always going to be cheaper and more energy-efficient to have dedicated cloud ha…
> It's widely understood that the big players are making profit on inference. This is most definitely not widely understood. We still don't know yet. There's tons of discussions about people disagreeing on whether it really is profitable. Unless you have proof, don't say "this is widely understood".
We need to see the cash flows.
Re: Gemma 4 on iPhone
#124Re: Gemma 4 on iPhone
#125Re: Gemma 4 on iPhone
#126Re: Gemma 4 on iPhone
#127Re: Gemma 4 on iPhone
#128Earlier quoted context omitted.
> or in the cloud but way more expensive then it is today. Why? It's widely understood that the big players are making profit on inference. The only reason they still have losses is because training is so expensive, but you need to do that no matter whether the models are running in the cloud or on your device. If you think about it, it's always going to be cheaper and more energy-efficient to have dedicated cloud ha…
> It's widely understood that the big players are making profit on inference. If you add in the cost of training, it’s not profitable. Not including the cost of training is a bit like saying the only cost of a cup of coffee is the paper cup it’s in. The only way OpenAI gets to charge for inference is by selling a product people can’t get elsewhere for much cheaper, which means billions in R&D costs. But because of co…
Obviously that doesn’t help them turn a profit, until they can stop growing training costs exponentially.
So it’s really a race to see whether growth in revenue or training costs decelerates first.
Re: Gemma 4 on iPhone
#129Nice! Tried on iPhone 16 pro with 30 TPS from Gemma-4-E2B-it model. Although the phone got considerably hot while inferencing. It’s quite an impressive performance and cannot wait to try it myself in one of my personal apps.
Still, absolutely fabulous. What a time to be alive!
Re: Gemma 4 on iPhone
#130> We collect information about your activity in our services