Live data from Hacker News

Apple's On-Device and Server Foundation Models

machinelearning.apple.com

291–300 of 562 posts

Re: Apple's On-Device and Server Foundation Models

#292
post #291

Earlier quoted context omitted.

[flagged]

Yes, it does. Should we keep arguing like on school playground?

Maybe, if you prefer. Honestly, I'm migrating a server fleet today, so notifications are hard to hear over these Apollo 6500 fans.

Re: Apple's On-Device and Server Foundation Models

#293
post #201

For people interested in AI research, there's nothing new here. IMO they should do a better job of referencing existing papers and techniques. The way they wrote about "adaptors" can make it seem like it's something novel, but it's actually just re-iterating vanilla LoRA. It was enough to convince one of the top-voted HackerNews comments that this was a "huge development". Benchmarks are nice though.

I think the thing they're saying that's novel, isn't what they have (LoRAs), but where and when and how they make them. Rather than just pre-baking static LoRAs to ship with the base model (e.g. one global "rewrite this in a friendly style" LoRA, etc), Apple seem to have chosen a bounded set of behaviors they want to implement as LoRAs — one for each "mode" they want their base model to operate in — and then set up a…

There is no way they would secretly train loras in the background of their user's phones. The benefits are small compared to the many potential problems. They describe some LoRA training infrastructure which is likely using the same capacity as they used to train the base models.

> ...each LoRA gets fine-tuned per user...

Apple would not implement these sophisticated user specific LoRA training techniques without mentioning them anywhere. No big player has done anything like this and Apple would want the credit for this innovation.

Re: Apple's On-Device and Server Foundation Models

#294

Earlier quoted context omitted.

I want to know what they consider "harmful". Is it going to refuse to operate for sex workers, murder mystery writers, or people who use knives?

They'll inject whatever ideology / dogma is "the current thing" into this.

[flagged]

Re: Apple's On-Device and Server Foundation Models

#295

Earlier quoted context omitted.

Those who dislike censorship and enjoy hacking avoid iPhones for obvious reasons.

People who understand cybersecurity hygiene use iPhones for obvious reasons

The reasons are very not obvious to me. Could you elaborate?

Re: Apple's On-Device and Server Foundation Models

#296
post #285

Earlier quoted context omitted.

You get considerably more ML FLOPS per dollar in a 4090 than any mac. It seems like the base M2 MAX is at roughly the same price point. It does grant you more RAM. Quadro and Tesla cards might be a different story. I would still like to see concrete FLOPS/$ numbers.

The M2 is a chip designed to be in a laptop (and it is quite powerful given its low power consumption). Presumedly they have a different chip or at least completely different configuration (RAM, network, etc.) in their data centers.

The interesting point here is that developers targeting the Mac can safely assume that the users will have a processor capable of significant AI/ML workloads. On the Windows (and Linux) side of things, there's no common platform, no assumption that the users will have an NPU or GPU capable of doing what you want. I think that's also why Microsoft was initially going for the ARM laptops, where they'd be sure that the required processing power is available.

Re: Apple's On-Device and Server Foundation Models

#297
post #286

Earlier quoted context omitted.

Bet it depends on the country. In the USA, you won't be able to ask about sex, but you can probably ask about tank man.

I would've thought the same until Microsoft started hiding tank man results in Bing. I'm not so sure if companies will start training different models for every oppressive regime.

Oh?

https://www.bing.com/images/search?q=tank%20man

Looks visible, to me.

Tiananmen Square even shows Tank Man on the first page, 13th and 15th entry, for me. Admittedly, I expected it more quickly on Tiananmen Square; but that might be because I, as a person, forgot that it's also a literal square with more stuff going on at it than a single moment in history.

Re: Apple's On-Device and Server Foundation Models

#298
post #254
post #149

Earlier quoted context omitted.

For the majority of the keynote they explicitly avoided the word AI instead substituting the word Intelligence, then Apple Intelligence, and then towards the end they said AI and ChatGPT once or twice. I think they saw the response to all the AI shoveling and Microsoft Recall and executed a fantastic strategy to reposition themselves in industry discussions. I still have tons of reservations about privacy and what th…

> makes me excited to develop for their platform in a way I haven't felt in a very, very, long time AI will ultimately do all the 'development', and will replace all apps. The integrations are going to be a temporary measure. Only apps that will survive are the ones that control things that apple cannot control (ie. how Uber controls its fleet)

Perhaps. It will be exciting to see if/how that happens. It does seem relatively far off still. At least some years.

Re: Apple's On-Device and Server Foundation Models

#299

For people interested in AI research, there's nothing new here. IMO they should do a better job of referencing existing papers and techniques. The way they wrote about "adaptors" can make it seem like it's something novel, but it's actually just re-iterating vanilla LoRA. It was enough to convince one of the top-voted HackerNews comments that this was a "huge development". Benchmarks are nice though.

Thing is, Apple takes these concepts and polishes them, makes them accessible to maybe not laypeople but definitely a much wider audience compared to those already "in the industry", so to speak.
Post reply on HN