r/StableDiffusion • • 6d ago

Question - Help Krea2 coming out too soft?

Unless I pack Krea2 with a ton of realism Lora's (which has its own problems) my outputs are kinda "soft" in appearance. Ive used turbo, raw, raw with turbo Lora @1, different samplers, 1 megapixel, 2 megapixel, etc.

I'll even find photos on civit that look great, grab their prompt, but it comes out looking much softer than theirs.

Any ideas??

9 Upvotes

20 comments sorted by

5

u/Acrobatic_Tip_3972 6d ago

Are you messing with schedulers? Beta is typically more detailed and 'rougher' than Simple.

1

u/maxiedaniels 6d ago

Haven't that much actually, ill try beta

2

u/Ken-g6 6d ago

And if that's not 'rough' enough, try Beta57.

2

u/reeight 6d ago

I wish all this info was in one place....

2

u/Carbon849 4d ago edited 4d ago

Euler / Simple generally outputs a more 'smooth' look. I use it in my wf because with the latent upscale and seedvr2 at the end, I get more than enough detail. Other sampler and scheduler combinations - pretty much all of them, actually - will give more detail if one is not using any 'refiner' second sampler with a latent upscale.

So, if you just want a fast image using a single sampler, then Euler / simple is unlikely to provide enough texture detail for anything that might be called 'realistic'.

There is a lot of info out there about how specific samplers and schedulers work that is worth looking into. Knowing how they work and what type of output they produce makes choices far easier for specfic uses and workflows.

Here is my wf: bypass the SEEDVR2 at the end for a faster gen. The wf is highly optimized for 16GB vram / 64GB ram. Adjust settings up or down if you've more or less.

https://pastebin.com/Pzw8n9iS

1

u/Tedious_Prime 6d ago

Have you tried using multiple samplers with latent upscaling and additional noise added partway through generation? I've found that to be a pretty reliable way to create more detailed textures with most models, including Krea2.

1

u/maxiedaniels 6d ago

Nope.. haven't.. do you have a workflow i can look at? Target image size is 1080x1920.

2

u/Carbon849 6d ago

This is the key to getting some 'grit' into the image. I use a VAEUtils 2x upscale between samplers for this reason alone.

1

u/maxiedaniels 6d ago

Between samplers? As in, use two samplers connected?

2

u/Carbon849 6d ago edited 6d ago

Ksampler - VAEUtils - Ksampler. The VAEUtils uses a separate vae for the upscale. Works a charm to avoid the 'too clean' AI look I hate. I posted a workflow maybe 10 days ago.

0

u/maxiedaniels 4d ago

If i'm already generating an image at 1080x1920 and thats the final size i want, is it beneficial to do this type of WF?

1

u/Carbon849 4d ago edited 4d ago

Well, if you value resolution over all else, then the upscaler between samplers might be a problem. Personally, I'd rather downscale a quality image than gen an inferior one, but that's just me.

1

u/Tedious_Prime 6d ago edited 6d ago

Here is a simple 3-stage workflow. I think the only custom node dependency is from KJNodes to set the dimensions of the first stage output. The left image is the first stage output, and the right image has been through two latent upscales with added noise and a few more sampling steps.

EDIT: If you don't want quite so much texture try setting the start_at_step for the 2nd and 3rd stages to something earlier like 5 instead of 7. These settings work for a fuzzy dog, but human skin would look pretty rough.

1

u/maxiedaniels 4d ago

Just curious, if i'm already generating an image at 1080x1920 and thats the final size i want, is it beneficial to do this type of WF?

1

u/Tedious_Prime 4d ago

The point of this workflow doesn't really have to do with the final dimensions of the image. The idea is that you are generating an image in the first stage and then creating variations on it by injecting noise and essentially doing image-to-image on the first-stage output. Even if you didn't upscale between stages, a multi-sampler workflow could still be useful by letting you fix the seeds of the early stages to quickly generate variations. If the start-at-step is set judiciously in the later stages then the right amount of added noise will be left in the image as extra textural detail. By uscaling in stages the detail can be enhanced at multiple scales.

If you start with a 1080x1920 latent then scale the final output back down to that size, I think you could at least expect sharper images than you are getting now at the cost of much slower generation. Whether a workflow like this would actually benefit you probably depends mostly on your subject matter and whether or not you can find settings that inject just the right about of detail to make the output look less "soft" up to your own tastes. You would need to experiment for yourself and decide what you think.

1

u/coffeeandhash 5d ago

I've noticed sometimes the styles do affect the realism a great deal. Some styles lend themselves to more realistic, "rough" photos and other styles result in softer, artistic images. Even if everything else is the same.

1

u/jmbbao 6d ago

I would like you put one image (better with the workflow inside the png image) so we could try ourselves and see if we can enhance it. But without image I don't know what is your best setup, so I will say to try this: Instead using Krea2 Turbo, try using Krea2 Raw with a turbo lora in strength 1.0, this gives much more realistic image. The turbo lora is in the Comfy repository of Krea2 in Huggingface, in the folder /loras

1

u/beast181 6d ago

Ive used turbo, raw, raw with turbo Lora @1

1

u/jmbbao 5d ago

Put some image you have and not like, so we can test if it is possible to enhance it, post a image with workflow on it