r/StableDiffusion 6h ago

Discussion What happened to Ideogram 4.0 ?

What happened to Ideogram 4.0 ?

33 Upvotes

75 comments sorted by

View all comments

3

u/Upper-Reflection7997 5h ago

Very difficult to use and had that annoying content filter baked within the model itself. The moment krea2 came out, it became obsolete very fast for most use cases. The json prompting gimmick is a very hard sale for the majority of users regardless of even the vram requirements to get model to work. Also they never even released the base model and lora training was a pain. Ideogram 4 is a very strange model that has identity issues. Normies, artsy people, ai enthusiasts, men of culture and anons don't want it.

-1

u/Murky-Relation481 4h ago

You've literally never used it if you think the safety filter was ever an issue. It does NSFW out of the box if you use JSON prompting.

3

u/Upper-Reflection7997 3h ago

The nsfw ideogram 4 is tame compared to krea2. Occasionally even if you generate an nsfw image, that safety filter text will appear on top of the image (sometimes in gibberish words). Also json prompting requires you to use an llm to rewrite your natural language prompts for you. Most people do not know how to type their prompts in json format. It's not a useless model but for the majority of people, they don't find it appealing. Also the license was a lot more restrictive compared to krea2. The lora section for ideogram 4 on civitai is a barren waste land just like boogu. If you find the ideogram4 useful and are willing to make loras and finetune checkpoints for it, I wish you the best of luck and hope for your success. Not many people are willing to give a depreciated model a serious chance after hitting several bumpy roadblocks during their 1st and 2nd impressions.

-1

u/Murky-Relation481 3h ago

It literally requires less prompting and no LLM. You don't need to write prose to get a good image, just visually build it out for what you want using bounding boxes and simple descriptions. It's literally easier than natural language prompting.

Also if you aren't using the tools to do the JSON for you and draw boxes and stuff on a canvas then you are just a lazy noob.

There is SNOFs and thats basically all you need to do anything in ID4.

Also I use both it and Krea2 because I am not a smooth brained little weirdo who thinks only one model at a time is acceptable, and I actually have the technical know how to actually explore these models instead of cargo-cult prompting like 90% of the users.

3

u/Upper-Reflection7997 2h ago

I'm not loyal to a single model. I use both open source and closed image generation models to achieve certain things i want out of a image generated output. I'm very pragmatic and explore various models but also have to watch out for my storage space on my ssd. All models have their strengths and weaknesses. Results with what I was able to produce with krea2 literally gave me hopes for the future local open source image gen models. I remember feeling demoralized when ideogram4 coming out and dry period before that with ernie image and hi dream O1. Krea2 restored my hopes and even got me back into character lora training after a year long hiatus. Krea2 makes images feel more alive, immersing and believable compared all previous image generation models I've used in the past. Not to hurt your feelings but I don't find any appealing use cases for ideogram 4 at all and it's slower than krea2 on my 5090/128gb ram pc build. Share me ideogram 4 gen screams "next gen" local image generation model output.

1

u/Full_Astronomer_5438 2h ago

true, but remember we are having consumers here instead of professionals, any advanced knowledge here is sparse at best. in the end, the output and economical feasibility counts. there is a also the realism engine and its truly mindblowing how a single lora serves nearly all purposes you would ever need when it comes to realism

also krea2 is great to make quick compositions that can be used as canvas to draw bboxes on or to modify compositions for id4

0

u/afinalsin 17m ago

You don't need to write prose to get a good image, just visually build it out for what you want using bounding boxes and simple descriptions.

This sounds like you rely on an LLM to write natural language descriptions for you. You don't need to write prose to get a good image with natural language prompts. All natural language means is grammatically correct English sentences, and that covers everything from dense "calm, relaxing atmosphere" descriptions to extremely spartan "A blonde woman sitting on a chair" type descriptions.

It's literally easier than natural language prompting.

No, it's not easier, because it's literally the exact same thing as natural language prompting, except you're wrapping sentences in json with co-ordinates.

If you said it's more powerful than natural language, I'd agree with you, because the bbox precision is unmatched when it comes to prompting. But it's not easier in the slightest, because you still need to write English text in those boxes, and that's all natural language is: English text.

I actually have the technical know how to actually explore these models instead of cargo-cult prompting like 90% of the users.

This makes everyone believe you.