Live data from Hacker News

MobileDiffusion: Rapid text-to-image generation on-device

blog.research.google

31–40 of 67 posts

Re: MobileDiffusion: Rapid text-to-image generation on-device

#32
post #30

So what could be some use cases for this apart from as a toy or for faking photos/art?

Generating memes to friends sparing search time. Describing sketches have them on the fly sharing with teams?

People search and remember things visually. Even if they're not consciously aware. So on the Cybershow [0] we decided to jump-in and use AI images as a quick way to visually tag episodes with something meaningful and fun.

We did that despite some moral ambivalence/uneasiness around AI "art".

For example, give me a "young and exciting Dana Meadows in front of a board of systems theory"

I'm not awful at photoshopping things, and sometimes that's the only way to get a specific image one has in mind. But it saves time and lets us concentrate on writing and researching instead.

TBH if an artist/illustrator came along and said "Let me do the episode icons even though you can't pay me yet" I'd feel inclined to ask the AI to step aside.

[0] https://cybershow.uk/episodes.php

Re: MobileDiffusion: Rapid text-to-image generation on-device

#33
post #10

Earlier quoted context omitted.

I'm interpreting it as they will be adding a layer of safety restrictions. Understandable given the furore of the recent Taylor Swift generated image incident. Everyone needs to do this and probably is already doing this. Search for "ChatGPT lobotomized" and you'll see plenty of complaints about the safety filters added by OpenAI.

I'm much more comfortable with the idea of AI watermarking images it creates instead of refusing to create images because of "safety", which in practice more often means not wanting to offend anyone. Imagine if word processors like Google Docs refused to write things you wanted to write because of mature themes. The important thing, in my opinion, is to make it a lot more difficult to pass off AI generated content as…

It being authentic or not isn't actually important in a lot of cases though. Consider someone like Mia Janin, who recently took her own life after been harassed using deepfakes. Everyone understood that the images weren't "authentic" but their power to cause distress was very real.

Re: MobileDiffusion: Rapid text-to-image generation on-device

#35
I never upvote any Google's A.I. research articles as most of the time it is: look what we have done, but we will never release anything.

OpenAi gets a lot of criticism for being closed, but at least I can play with their api most of the time.

What's the point of this if we will never be able to use this?

Re: MobileDiffusion: Rapid text-to-image generation on-device

#37
post #28

Earlier quoted context omitted.

Not entirely true either. If it thinks it has network but it's flakey, it won't translate offline, it will say there is network error and will give you a button to retry. No button to do offline. Additionally, in airplane mode it heavily doesn't want to translate, in my use case I have to go to saved translations as otherwise it won't even let me type what I need to translate.

I just tried airplane mode on my pixel 7 pro and it seemed to be able to translate from the camera without problems It doesn't seem to do it "live" in the preview without network access, though. And the translation app seemed to get into a bad state and fail to download the language packs without first clearing the data, saying I need to download the pack, but the language list showing it already was. Though I haven'…

Pixel 8 / Tensor G3 specifically is rumored to have a performance or heat issue. The scope of that rumored issue is limited to that generation.

Re: MobileDiffusion: Rapid text-to-image generation on-device

#38
post #18

Earlier quoted context omitted.

All the Pixel AI stuff runs on the cloud anyway. Just try using it in airplane mode.

not entirely true. translation is done on device

Magic eraser in google photos is also on device. (I have not given Google Photos Network permission and don't have google play services installed)

The voice recorder transcription is also on device, but I haven't gotten it to work without google play services on GrapheneOS

Re: MobileDiffusion: Rapid text-to-image generation on-device

#39
post #35

I never upvote any Google's A.I. research articles as most of the time it is: look what we have done, but we will never release anything. OpenAi gets a lot of criticism for being closed, but at least I can play with their api most of the time. What's the point of this if we will never be able to use this?

Corporate AND personal marketing.

AI researchers can make any claim, the risk of getting busted is close to none.

Didn't work ? Well dataset was different

Didn't work ? Well code was different

Can I try your work ? Well it's proprietary / I don't have access / We shutdowned the cluster

But the result is guaranteed increase in salary and job opportunities.

Since these companies are publicly listed, they are by definition encouraged and encouraging to make grandiose claims in order to make themselves more attractive to investors, and they can blame the individuals if it becomes discovered.

My favorite being that Bard (PaLM version) is sentient, but it was too big this time.

Imagine a large pharmaceutical company claiming they can cure very important diseases, but the results cannot be independently verified, nor audited.

It’s ok in the short-term, but not when you make that claim during few years.

Re: MobileDiffusion: Rapid text-to-image generation on-device

#40

Earlier quoted context omitted.

I'm much more comfortable with the idea of AI watermarking images it creates instead of refusing to create images because of "safety", which in practice more often means not wanting to offend anyone. Imagine if word processors like Google Docs refused to write things you wanted to write because of mature themes. The important thing, in my opinion, is to make it a lot more difficult to pass off AI generated content as…

It being authentic or not isn't actually important in a lot of cases though. Consider someone like Mia Janin, who recently took her own life after been harassed using deepfakes. Everyone understood that the images weren't "authentic" but their power to cause distress was very real.

Is there a difference between being harassed with deepfakes vs fakes?

Photoshop has the power to cause distress too when used maliciously.

You go after the aggressors, not the tool used for aggression.

Post reply on HN