Live data from Hacker News

MobileDiffusion: Rapid text-to-image generation on-device

blog.research.google

21–30 of 67 posts

Re: MobileDiffusion: Rapid text-to-image generation on-device

#22
post #11
post #6

Google has fallen so far. Both Inception and Mobilenet were released openly and changed the entire AI world. Nowadays we just get blog posts about results that were supposedly achieved, an accompanying paper that can’t be reproduced (because of Google’s magical “private datasets”), and some screencaps of a cool application of the tech that is virtually guaranteed to never make it to product.

Probably because the actual product is garbage. Remember the Google assistant demo, where it booked a table at a restaurant? That never materialized. Google assistant is just eating crayons today.

It did materialize.

https://www.reddit.com/r/googlehome/comments/ezv3us/google_a...

Or the comments under https://youtu.be/-RHG5DFAjp8

It's probably hard to trigger these days because most places support OpenTable or similar.

Re: MobileDiffusion: Rapid text-to-image generation on-device

#24

Kind of funny that they show the iphone 15 pro and the Samsung S24 in the comparison chart, but not their own phone the google pixel 8. (I know it will perform worse than both phones)

There is some truth to it.

I don’t see it as a disadvantage, since Google markets services on both devices you mentioned. Hardly anyone will abandon its iPhone in favor of a Pixel just for a Google service.

So I think it’s ok what Google did marketing wise.

Re: MobileDiffusion: Rapid text-to-image generation on-device

#25
post #23

[flagged]

What a sadly narrow definition of the hacker spirit you seem to have.

I enjoy reading articles about interesting research, regardless if there's a practical application of it or not.

(If you don't like this sort of thing, just flag it and move on. No need to waste comment space with denouncements.)

Re: MobileDiffusion: Rapid text-to-image generation on-device

#26
post #9

Earlier quoted context omitted.

I would interpret it as "expect to see this powering some features in the next-generation Pixel".

We've already seen this progression - they debuted Magic Eraser as a cloud feature, then with the Pixel 8 they got it running locally on the device. But they also introduced Magic Editor with the Pixel 8, running on the cloud, and the Pixel 9 or 10 will probably run it on-device.

it may turn out more like the imagen timeline

2022-05 - google imagen research paper posted https://news.ycombinator.com/item?id=31484562

2022-12 - imagen developers leave google to form ideogram

2023-08 - ideogram ships a version of imagen, free, for anyone who wants to use it https://ideogram.ai/publicly-available

2023-12 - google "imagen 2" is officially "generally available for Vertex AI customers on the allowlist (i.e., approved for access)." https://news.ycombinator.com/item?id=38628417

Re: MobileDiffusion: Rapid text-to-image generation on-device

#27
post #10

> With superior efficiency in terms of latency and size, MobileDiffusion has the potential to be a very friendly option for mobile deployments given its capability to enable a rapid image generation experience while typing text prompts. And we will ensure any application of this technology will be in-line with Google’s responsible AI practices. So I'm interpreting this that it won't ever get released.

I'm interpreting it as they will be adding a layer of safety restrictions. Understandable given the furore of the recent Taylor Swift generated image incident. Everyone needs to do this and probably is already doing this. Search for "ChatGPT lobotomized" and you'll see plenty of complaints about the safety filters added by OpenAI.

I'm much more comfortable with the idea of AI watermarking images it creates instead of refusing to create images because of "safety", which in practice more often means not wanting to offend anyone. Imagine if word processors like Google Docs refused to write things you wanted to write because of mature themes. The important thing, in my opinion, is to make it a lot more difficult to pass off AI generated content as being authentic and to make provenance traceable if you were to do something like create revenge porn with AI, but not to make AI refuse to create explicit material at all.

Re: MobileDiffusion: Rapid text-to-image generation on-device

#28
post #18

Earlier quoted context omitted.

not entirely true. translation is done on device

Not entirely true either. If it thinks it has network but it's flakey, it won't translate offline, it will say there is network error and will give you a button to retry. No button to do offline. Additionally, in airplane mode it heavily doesn't want to translate, in my use case I have to go to saved translations as otherwise it won't even let me type what I need to translate.

I just tried airplane mode on my pixel 7 pro and it seemed to be able to translate from the camera without problems

It doesn't seem to do it "live" in the preview without network access, though.

And the translation app seemed to get into a bad state and fail to download the language packs without first clearing the data, saying I need to download the pack, but the language list showing it already was. Though I haven't even opened it since I transferred it from my old phone, so if there's some phone-specific stuff going on that might have got messed up.

Re: MobileDiffusion: Rapid text-to-image generation on-device

#29
post #25
post #23

[flagged]

What a sadly narrow definition of the hacker spirit you seem to have. I enjoy reading articles about interesting research, regardless if there's a practical application of it or not. (If you don't like this sort of thing, just flag it and move on. No need to waste comment space with denouncements.)

If you don't like their comment, just flag it and move on. No need to waste comment space with more denouncements.
Post reply on HN