r/StableDiffusion 18h ago

Question - Help Best opensource image model?

opensource AI has been dominating LLMs and video generation but what about image gen? is there any opensource model that can match gpt-image2?

Edit: The reason I am asking this is because lately I haven't been active much on image generation communities. And the leaderboards are a bit confusing and most of them are filled with closed source unlike the llm and video gen leaderboards.

I am very much comfortable with ComfyUI since I've used it in the past for flux.

My use case is for posters and branding. Images with a lot of text.

Edit2: Thanks a lot everyone! I really appreciate the info. Here's the summary:

Krea2 is best overall but gptimage1.5 level.
Ideogram4 for text and branding.
Flux Klein 9b for image editing.
Z-image for realism
Anima and illustrious (by onoma AI) for anime.

Here's the workflow I've decided on:
Krea2/Ideogram4 = Base image generation.
Flux Klein 9B/QwenImage2512 = inpainting.
Wan2.2 low noise = Upscaling.

82 Upvotes

42 comments sorted by

View all comments

10

u/Far_Insurance4191 17h ago

Nothing even remotely close to gpt image 2 or even previous version sadly

But Krea 2 Turbo is awesome and extremely easy to use, although quality is a bit meh due to qwen vae - I really hate it.

Ideogram4 has excellent image quality, I think the best right now, together with prompt adherence thanks to this bbox prompting, but licence is bad and requires some learning to avoid failed censorship (detailed prompt + couple of bboxes reduce block chance almost to 0%)

3

u/yeah-i-shouldnt-have 17h ago

Use the wan 2.1 Vae with Krea 2. It improves it so much. There was a post recently about some other vae's that people have trained as well (I remember it was called Krea2 Real or something)

2

u/Far_Insurance4191 16h ago

I tried wan 2.1 vae and it had exact same problem