Live data from Hacker News

Comparing Adobe Firefly, Dalle-2, and OpenJourney

blog.usmanity.com

121–130 of 139 posts

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#121

Earlier quoted context omitted.

Jaysus. I'm going to sound like an entitled whiny old guy shouting at clouds, but - what the hell; with all the knowledge being either locked and churned on Discord, or released in form of YouTube videos with no transcript and extremely low content density - how is anyone with a job supposed to keep up with this? Or is that a new form of gatekeeping - if you can't afford to burn a lot of time and attention as if in s…

The difference being that youtube videos can make more money for the author. Anyway, it's all open source, so feel free to make a wiki

I would if I could keep up with the videos :).

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#122

Amazing how quickly Dalle-2 went from among the best image transformers to among the worst.

It might be a case of them seeing way more potential with LLMs compared to image generation.

It’s more that their moat got obliterated on image gen.

If stable diffusion didn’t launch Dall-e 2 would have been still valuable.

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#123

Earlier quoted context omitted.

The difference being that youtube videos can make more money for the author. Anyway, it's all open source, so feel free to make a wiki

I would if I could keep up with the videos :).

I think it'd have been convenient for me as well if the AI tool that has access to YouTube videos would've been able to answer queries . But it takes 5 minutes to reply and I forgot it's name. It was on the front page recently

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#124

Earlier quoted context omitted.

How does anyone keep up with anything? It's a visual thing. A lot of people are learning drawing, modeling, animation etc in the exact same way - by watching YouTube (a bit) and experimenting (a lot).

Picking images from generated sets is a visual thing. Tweaking ControlNet might be too (IDK, I've never got a chance to use it - partly because of what I'm whining about here). However, writing prompts, fine-tuning models, assembling pipelines, renting GPUs, figuring out which software to use for what, where to get the weights, etc. - none of this is visual. It's pretty much programming and devops. I can't see how co…

There's a relatively thin layer between the papers and implementations, which is another way of saying this stuff is still for researchers and assumes a requisite level of background with them. It sounds like you'd benefit from seeking out the first party sources.

This is where video demonstrations come in handy. Since many concepts are novel, it's uncommon to find anyone who deeply understands them, but it's very easy to find people who have picked up on some tricks of the interfaces, which they're happy to click through. I think gradio/automatic1111 makes learning harder than it needs to be by hiding what it's doing behind its UI, while eg- comfyui has a higher initial learning curve but provides a more representational view of process and pipelines.

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#125

For reference, here's what you can get with a properly tweaked Stable Diffusion, all running locally on my PC. Can be set up on almost any PC with a mid range GPU in a few minutes if you know what you're doing. I didn't do any cherry picking; this is the first thing it generated. 4 images per prompt. 1st prompt: https://i.postimg.cc/T3nZ9bQy/1st.png 2nd prompt: https://i.postimg.cc/XNFm3dSs/2nd.png 3rd prompt: https:…

You got incorporated into the article! Nice.

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#126

Earlier quoted context omitted.

Yet you are posting this in a thread where GP provided actual examples of the opposite. Look for another comment above/below, there are MJ-generated samples which are comparable but also less coherent than the result from a much smaller SD model. And in case of MJ hallucinations cannot be fixed. MJ is good but it isn't magic, it just provides quick results with little experience required; prompt understanding is stil…

>Yet you are posting this in a thread where GP provided actual examples of the opposite. Opposite of what ? OP posts results from a tuned model.

If course it's a tuned model. Why would anyone use stock SD these days?

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#127
post #39

Earlier quoted context omitted.

Can you elaborate on “properly tweaked”? When I use one of the Stable Diffusion and AUTOMATIC1111 templates on runpod.io, the results are absolutely worthless. This is using some of the popular prompts you can find on sites like prompthero that show amazing examples. It’s been serious expectation vs. reality disappointment for me and so I just pay the MidJourney or DALL-E fees.

You're not going to get even close to Midjourney or even Bing quality on SD without finetuning. It's that simple. When you do finetune, it will be restricted to that aesthetic and you won't get the same prompt understanding or adherence. For all the promise of control and customization SD boasts, Midjourney beats it hands down in sheer quality. There's a reason like 99% of ai art comic creators stick to Midjourney de…

Midjourney has a riduculously restrictive keyword filter. You should have mentioned that.

Also I see nothing wrong with using different models for different purposes.

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#128
post #87

Earlier quoted context omitted.

I'm using my own custom trained model. Here, I've uploaded it to civitai: https://civitai.com/models/94176 There are plenty of other good models too though.

Any tips or guides you followed on training your custom model? I've done a few LoRAs and TI but haven't gotten to my own models yet. Your results look great and I'd love a little insight into how you arrived there and what methods/tools you used.

Make sure that you have enough vram. I can train loras with 8 gb easily, but when I tried to train a model - it gives me an oom error.

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#129
post #28

Earlier quoted context omitted.

I'm presuming you're not including Stable Diffusion when you say this; the fact that SD and its variants are defacto extremely "free and open source" presently put it way ahead of anything else, and are likely to do so for some time.

As far as I can tell anyone who’s creating images is using midjourney. This is likely the same “Linux is open so it’s way better” tell that to the trillion dollar companies that bet against that.

[deleted]

Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney

#130
post #111

Earlier quoted context omitted.

I find it curious because (a) if they don't care about text2image, why launch it as a service to begin with? (b) if they don't care now, why keep it up and let it keep consuming resources, human & GPU? (c) if they do still care, because as other models & services have demonstrated there's a ton of interest in text2image, why not invest the relatively minor amount of resources it would take to keep it competitive (loo…

Maybe they keep it up just so that they have something in txt2img space? It may not be the best, or even good, but you don't know that until you try it, and until then, it just enhances the value of the OpenAI platform. E.g. if you're building something backed by OpenAI LLMs, and are thinking about future txt2img integration, the existence of Dall-E might stop you from "shopping around" txt2img services in advance. T…

Why do they need to have something in text2image? It in no way builds lockin to the API or anything, especially with how gimped it is.

1. Yes, they are. Look at the constant iterative rollouts of GPTs 2. Most of which is useless to them, not that they have made any use of it 3. the fact that it would be so easy to improve, and they haven't, only emphasizes my point. 4. sure, that could be useful. Except there's zero integration or mention. (They haven't even opened up the vision part of GPT-4 yet.) 5. the fact that it would be so easy to improve, and they haven't, only emphasizes my point. 6. why wait for GPT-5 possibly years from now?

Post reply on HN