Dall-E 2
311–320 of 511 posts
Re: Dall-E 2
#312Sam Altman demonstrates Dall-E 2 using twitter suggestions - https://news.ycombinator.com/item?id=30933478 - April 2022 (3 comments)
Re: Dall-E 2
#313The timing of the Dall-E 2 launch an hour ago seems to correspond with a recent piece of investigative journalism by Buzzfeed News about one of Sam Altman's other ventures, published 15 hours ago and discussed elsewhere actively on HN right now: https://news.ycombinator.com/item?id=30931614 I point this out because while Dall-E 2 seems interesting (I'm out of my depth, so delegating to the conversation taking place h…
I listed some of them here - https://news.ycombinator.com/item?id=30934732, just because I remembered there had been previous discussions and listing related previous discussions is a thing.
Re: Dall-E 2
#314Earlier quoted context omitted.
How do you run such a Google Colab thing? I don't see a run button? On.. maybe "Runtime -> Run All" from the menu ... Shows me a spinning circle around "Download model" ... 26% ... Fascinating, that Google offers you a computer in the cloud for free .. Now it is running the model. Wow, I'm curious .. Ha, it worked! Nothing compared to the images in the Dall-E 2 article but still impressive.
Google is a company with a lot of spare VMs and GPUs. However, the free GPU is now a K80 which is obsolete and barely sufficient for running these types of models.
Re: Dall-E 2
#315Re: Dall-E 2
#316I'm only part way through the paper, but what struck me as interesting so far is this: In other text-to-image algorithms I'm familiar with (the ones you'll typically see passed around as colab notebooks that people post outputs from on Twitter), the basic idea is to encode the text, and then try to make an image that maximally matches that text encoding. But this maximization often leads to artifacts - if you ask for…
I'm not sure if I'm speaking clearly, I just don't understand, what's the difference between training "text encoding to an image" vs "text embedding to image embedding". In both cases you have some kind of "sunset" (even though it's obviously just a dot in a multi-dimension space, not the letters) on the left, and you try to maximize it when training the model to get either a image-embedding or a image straight away.
Re: Dall-E 2
#317I'm only part way through the paper, but what struck me as interesting so far is this: In other text-to-image algorithms I'm familiar with (the ones you'll typically see passed around as colab notebooks that people post outputs from on Twitter), the basic idea is to encode the text, and then try to make an image that maximally matches that text encoding. But this maximization often leads to artifacts - if you ask for…
Do you think some of these techniques could be slightly modified, and applied to DNA sequences?
Re: Dall-E 2
#318Earlier quoted context omitted.
Literally everyone on this website is in denial. They all approach it by asking which fields will be safe. No field is safe. “But it’s not going to happen for a long time.” Climate deniers say the same thing and you think they should be wearing the dunce hat? The average person complains bitterly about climate deniers who say that it’s “my grandkids problem lol” but when I corner the average person into admitting AI…
I'm trying to understand your point, because I think I agree with you, but it's covered in so much hyperbole and invective I'm having a hard time getting there. Can you scale it back a little and explain to me what you mean? Something like: AI is going to replace jobs at such scale that our current job-based economic system will collapse?
Re: Dall-E 2
#319It's becoming clear that efficient work in the future will hinge upon one's ability to accurately describe what one wants . Unpacking that -- a large piece is the ability to understand all the possible "pitfalls" and "misunderstandings" that could happen on the way to a shared understanding. While technical work will always have a place -- I think that much creative work will become more like the management of a team…
You can definitely make them incremental. You can give it a task like "make a more accurate description from initial description and clarification". Even GPT-3-based models available today can do these tasks.
Once this is properly productionized it would be possible to implement stuff just talking with a computer.