Like many others, it seems, I have also been blown away by DALL-E 2. When I got access on Sunday, I first tried a lot of different prompts and got some interesting results. One semirandom one, “A photograph of a professor playing a grand piano on a rainy night in Tokyo,” produced some very atmospheric images. I then went down a rabbit hole of variations on that prompt (“A painting of...,” “A line drawing of...,” “A p…
DALL·E 2 prompt book [pdf]
91–100 of 157 posts
Re: DALL·E 2 prompt book [pdf]
#92Prompt crafting is quickly becoming an art. I just found out yesterday that there's actually market places for buying and selling prompts [0]. It can really make a big difference if you can tune the image by adding the right words. Midjourney [1] even allows things such as adjusting the weight of each keyword or how "literal" the AI should take your prompt. [0] https://promptbase.com/ [1] https://midjourney.gitbook.i…
Does anyone know what kind of prompt generates such clean, consistent results like these? https://promptbase.com/prompt/clay-emojis https://promptbase.com/prompt/polygon-animals
Re: DALL·E 2 prompt book [pdf]
#93Is DALL-E different then other models for example ones on hugging face? Or is it relatively the same just trained to a ridiculous amount and that's why it's results are so good?
I love Dall-E but I still use other models. Some of my favourite results have come from JAX CLIP Guided Diffusion:
https://colab.research.google.com/drive/12Bod44YVIXYRh39WRqp...
Disco Diffusion still holds up for painterly stuff. Majesty is great for portaits. MidJourney can beat/match Dall-E for lots of styles. I got very good results from Multi-Perceptor VQGAN+CLIP v4 for matching artist styles.
etc etc.
Dall-E is amazing and versatile but it's often lacking some "soul" that I get from other models.
Re: DALL·E 2 prompt book [pdf]
#94If you post those images online, they seem to ban you.
Re: DALL·E 2 prompt book [pdf]
#95Re: DALL·E 2 prompt book [pdf]
#96If you right click -> "save image as" on openAI, the image will be saved without their logo in the corner (it's done as some kind of CSS overlay). If you post those images online, they seem to ban you.
Re: DALL·E 2 prompt book [pdf]
#97Earlier quoted context omitted.
How long did it take for you to generate your images? I've been using https://www.craiyon.com/ for fun but the wait times always results in me getting distracted elsewhere.
I've been using Midjourney [1] (it's not free, 25 photos demo, then 10$ for ~200 photos or 30$ for unlimited I think?). It's fairly fast, ~20s for a grid of 4, then as much for upscaling. I like the controls, it lets you do variations and tweak the image as you go. It's not as good for doing concrete asks, but it's very good for getting specific vibes. The website feed [1] requires Discord login to view examples, but…
* $10 for ~200 prompts * $30 for ~900 prompts + unlimited "slow" prompts where your job is put at the end of the queue and you have to wait longer (no idea how much longer though.. are we talking about seconds or hours here?)
Re: DALL·E 2 prompt book [pdf]
#98If you right click -> "save image as" on openAI, the image will be saved without their logo in the corner (it's done as some kind of CSS overlay). If you post those images online, they seem to ban you.
Who bans you? Instagram? Reddit?
Re: DALL·E 2 prompt book [pdf]
#99I just got access to the DALL-E 2 beta, and it's a ton of fun to make pictures out of everyday occurrences as prompts. Someone else here on HN observed that everyday people don't "get" how huge this all is. I experimented with asking random acquaintances at a local cafe for prompts and showed them the generated pictures. All but one person was totally unimpressed. If everything feels like magic, then what's one more…
Re: DALL·E 2 prompt book [pdf]
#100In my mind, the main eras of content on the internet look something like this:
Epoch 1: Pure, unblemished user generated content. Message boards and forums rule.
Epoch 2: More user generated content + a healthy mix of recycled user generated content. e.g. Reddit.
Epoch 3 (Now): Fake user generated content (limits to how much because humans still have to generate it). e.g. Amazon reviews, Cambridge Analytica.
Epoch 4: Advanced generative models means (essentially) zero friction for creating picture and text content. GPT3, Dalle-2.
Epoch 5: Generative models for videos, game over.
IMO, the future of the internet feels like a totally disastrous (un)reality. If addictive content recommended by the likes of TikTok has proven anything, it's that users ultimately don't care _what_ the content is, as long as it keeps their attention. It doesn't matter if it comes from a human or a machine. The difference is that in a world where the marginal cost of generating content is essentially zero, that content can and will be created and manipulated by large malicious actors to sway public opinion.
The Dead Internet Theory will fast become reality. This terrifies me.
[1] https://www.theatlantic.com/technology/archive/2021/08/dead-...