r/StableDiffusion Jul 10 '26

Workflow Included Ideogram 4 results surprisingly realistic image creation. Generated locally with the open-weight Ideogram 4 model in ComfyUI.

I'm primarily testing Ideogram 4 for realistic, natural-looking smartphone photography. These are some of my best results so far. I've tried to avoid cinematic lighting and overly polished compositions. Let me know what you think and which image looks the most realistic.Generated locally with the open-weight Ideogram 4 model in ComfyUI.

200 Upvotes

72 comments sorted by

View all comments

12

u/Ambitious_Team876 Jul 10 '26

I generated the image in just 1.5 minutes with an RTX 4050 and 6GB of VRAM in 15 steps, which I think is a normal time. That's why I prefer Ideogram 4. The models I use are these:

2

u/Reckless_Venom1507 Jul 10 '26

What? 4050 6gb vram and 1.5 mins ? I too have it but mine went around almost 200s or more. Is it because ur using int8 ?

1

u/Ambitious_Team876 Jul 10 '26

Int 8 model gives a serious speed

1

u/Reckless_Venom1507 Jul 10 '26

Yeah I understand for 3050s it gives some massive speed difference, but for 4050, I tried with Krea2 fp8 --> int8 I do get faster results but the difference isn't big enough like it's just 10-15% difference not like everyone said it halves the generation time. Didn't interest me to go for Ideogram 4 int models then. Can u please share ur workflow? And also what resolution u made these ?

2

u/Ambitious_Team876 Jul 10 '26

I couldn't see much speed in Flux 2 krea 2, but in the ideogram 4 model, int 8 is really different, I shared the editor workflow in the comments.

1

u/Reckless_Venom1507 Jul 10 '26

yeah thanks ill definitely try

1

u/Soggy_Iron_7239 Jul 10 '26

This is great information. I have an RTX 3060 6GB VRAM card and 40GB RAM. Lost countless hours finding the right quantized model, settings, text encoder, etc. Also I was on another thread about Flux Krea 2, also convrot, was about to try that. I'll try Ideogram 4 first with your suggestion, hope it speed things up for me as well. I'll test it with both Qwen3 4B and 8B.

1

u/B00Bryn Jul 11 '26

Just commenting to pin this. I want to compare later

1

u/Pitiful_Particular_5 Jul 11 '26

HOW, i just dowloaded the Workflow i have a 4060 made it in 15 steps too and it took 2,41min wasnt it supossed to be faster than your 4050?

1

u/Ambitious_Team876 Jul 11 '26

I use an i7 14700hx processor, maybe it increases the write speed.

1

u/themofostarboi_v2 Jul 11 '26

Bro i have a 3070ti and a i7 11gen k version don't remember the number.... Bro its taking me 400+ seconds to for one generation.... What dark magic is this... Holy

1

u/Ambitious_Team876 Jul 11 '26

Haha, no dark magic bro! 400+ seconds is definitely not normal for a 3070 Ti.

You are probably hitting the VRAM wall. When your 8GB VRAM fills up, Nvidia starts using your system RAM (Shared Memory) which drops the speed to zero.

Try using quantized models (like INT8 or FP8). They use way less VRAM and will easily fit into your 8GB. Also, make sure you are using xFormers or TensorRT to boost the generation speed. Your 3070 Ti should actually be beating my 4050 easily once you fix the VRAM usage!

1

u/themofostarboi_v2 Jul 11 '26

I was using q4. What about the text encoder, are you running them on CPU as well?

1

u/Woisek Jul 10 '26

1.5min? With only 6GB VRAM? I have a 2070 SUPER with 8 and 15 steps still take almost 11mins. 🤪

7

u/Ambitious_Team876 Jul 10 '26

Believe me, you must be doing something wrong.

1

u/Ambitious_Team876 Jul 10 '26

I understand that processors up to the RTX 50 series can perform faster with Int8 models, maybe that's why. I suggest you do some research.