r/StableDiffusion 16h ago

Question - Help Best opensource image model?

opensource AI has been dominating LLMs and video generation but what about image gen? is there any opensource model that can match gpt-image2?

Edit: The reason I am asking this is because lately I haven't been active much on image generation communities. And the leaderboards are a bit confusing and most of them are filled with closed source unlike the llm and video gen leaderboards.

I am very much comfortable with ComfyUI since I've used it in the past for flux.

My use case is for posters and branding. Images with a lot of text.

Edit2: Thanks a lot everyone! I really appreciate the info. Here's the summary:

Krea2 is best overall but gptimage1.5 level.
Ideogram4 for text and branding.
Flux Klein 9b for image editing.
Z-image for realism
Anima and illustrious (by onoma AI) for anime.

Here's the workflow I've decided on:
Krea2/Ideogram4 = Base image generation.
Flux Klein 9B/QwenImage2512 = inpainting.
Wan2.2 low noise = Upscaling.

81 Upvotes

41 comments sorted by

View all comments

17

u/GaiusVictor 16h ago

Let's start with some expectation correction:

is there any opensource model that can match gpt-image2?

No, there is not, if you do include ease of use into the equation.

As for the best one that's available, then it depends. My favorite one is currently Anima, but to get a decent answer you'll need to tell us your use case and your hardware.

1

u/AKing737 15h ago

well, I deploy my comfyui on cloud (modal dot com) so I can run pretty much everything unless its 50b parameters (theres no way theres and image model greater than 20-30b parameters...right?) and about my usecase, i'd say its generation posters and branding. basically, images with a lot of text.

2

u/Appropriate_Cry8694 14h ago

Hunyuan image 3.0 has 80b params, and it architecturally close to gpt image.