Earlier quoted context omitted.
Jaysus. I'm going to sound like an entitled whiny old guy shouting at clouds, but - what the hell; with all the knowledge being either locked and churned on Discord, or released in form of YouTube videos with no transcript and extremely low content density - how is anyone with a job supposed to keep up with this? Or is that a new form of gatekeeping - if you can't afford to burn a lot of time and attention as if in s…
The difference being that youtube videos can make more money for the author. Anyway, it's all open source, so feel free to make a wiki
Comparing Adobe Firefly, Dalle-2, and OpenJourney
121–130 of 139 posts
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#122Amazing how quickly Dalle-2 went from among the best image transformers to among the worst.
It might be a case of them seeing way more potential with LLMs compared to image generation.
If stable diffusion didn’t launch Dall-e 2 would have been still valuable.
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#123Earlier quoted context omitted.
The difference being that youtube videos can make more money for the author. Anyway, it's all open source, so feel free to make a wiki
I would if I could keep up with the videos :).
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#124Earlier quoted context omitted.
How does anyone keep up with anything? It's a visual thing. A lot of people are learning drawing, modeling, animation etc in the exact same way - by watching YouTube (a bit) and experimenting (a lot).
Picking images from generated sets is a visual thing. Tweaking ControlNet might be too (IDK, I've never got a chance to use it - partly because of what I'm whining about here). However, writing prompts, fine-tuning models, assembling pipelines, renting GPUs, figuring out which software to use for what, where to get the weights, etc. - none of this is visual. It's pretty much programming and devops. I can't see how co…
This is where video demonstrations come in handy. Since many concepts are novel, it's uncommon to find anyone who deeply understands them, but it's very easy to find people who have picked up on some tricks of the interfaces, which they're happy to click through. I think gradio/automatic1111 makes learning harder than it needs to be by hiding what it's doing behind its UI, while eg- comfyui has a higher initial learning curve but provides a more representational view of process and pipelines.
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#125For reference, here's what you can get with a properly tweaked Stable Diffusion, all running locally on my PC. Can be set up on almost any PC with a mid range GPU in a few minutes if you know what you're doing. I didn't do any cherry picking; this is the first thing it generated. 4 images per prompt. 1st prompt: https://i.postimg.cc/T3nZ9bQy/1st.png 2nd prompt: https://i.postimg.cc/XNFm3dSs/2nd.png 3rd prompt: https:…
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#126Earlier quoted context omitted.
Yet you are posting this in a thread where GP provided actual examples of the opposite. Look for another comment above/below, there are MJ-generated samples which are comparable but also less coherent than the result from a much smaller SD model. And in case of MJ hallucinations cannot be fixed. MJ is good but it isn't magic, it just provides quick results with little experience required; prompt understanding is stil…
>Yet you are posting this in a thread where GP provided actual examples of the opposite. Opposite of what ? OP posts results from a tuned model.
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#127Earlier quoted context omitted.
Can you elaborate on “properly tweaked”? When I use one of the Stable Diffusion and AUTOMATIC1111 templates on runpod.io, the results are absolutely worthless. This is using some of the popular prompts you can find on sites like prompthero that show amazing examples. It’s been serious expectation vs. reality disappointment for me and so I just pay the MidJourney or DALL-E fees.
You're not going to get even close to Midjourney or even Bing quality on SD without finetuning. It's that simple. When you do finetune, it will be restricted to that aesthetic and you won't get the same prompt understanding or adherence. For all the promise of control and customization SD boasts, Midjourney beats it hands down in sheer quality. There's a reason like 99% of ai art comic creators stick to Midjourney de…
Also I see nothing wrong with using different models for different purposes.
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#128Earlier quoted context omitted.
I'm using my own custom trained model. Here, I've uploaded it to civitai: https://civitai.com/models/94176 There are plenty of other good models too though.
Any tips or guides you followed on training your custom model? I've done a few LoRAs and TI but haven't gotten to my own models yet. Your results look great and I'd love a little insight into how you arrived there and what methods/tools you used.
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#129Earlier quoted context omitted.
I'm presuming you're not including Stable Diffusion when you say this; the fact that SD and its variants are defacto extremely "free and open source" presently put it way ahead of anything else, and are likely to do so for some time.
As far as I can tell anyone who’s creating images is using midjourney. This is likely the same “Linux is open so it’s way better” tell that to the trillion dollar companies that bet against that.
Re: Comparing Adobe Firefly, Dalle-2, and OpenJourney
#130Earlier quoted context omitted.
I find it curious because (a) if they don't care about text2image, why launch it as a service to begin with? (b) if they don't care now, why keep it up and let it keep consuming resources, human & GPU? (c) if they do still care, because as other models & services have demonstrated there's a ton of interest in text2image, why not invest the relatively minor amount of resources it would take to keep it competitive (loo…
Maybe they keep it up just so that they have something in txt2img space? It may not be the best, or even good, but you don't know that until you try it, and until then, it just enhances the value of the OpenAI platform. E.g. if you're building something backed by OpenAI LLMs, and are thinking about future txt2img integration, the existence of Dall-E might stop you from "shopping around" txt2img services in advance. T…
1. Yes, they are. Look at the constant iterative rollouts of GPTs 2. Most of which is useless to them, not that they have made any use of it 3. the fact that it would be so easy to improve, and they haven't, only emphasizes my point. 4. sure, that could be useful. Except there's zero integration or mention. (They haven't even opened up the vision part of GPT-4 yet.) 5. the fact that it would be so easy to improve, and they haven't, only emphasizes my point. 6. why wait for GPT-5 possibly years from now?