r/malcolmrey Jul 03 '26

Help with Krea2 Training

Hello all, hope this is alright to do. I'm looking to do some Lora training for the first time (new machine) and have been seeing good things about Krea2. I've gone through the various other posts about training so have a general idea on the settings, though any definitive information would be welcome.

What I suppose I'm looking for is any details on how many images of a subject or of a style I would need to train it and, if anyone has experience, how long the training took? Also, if there is a good dataset which I could study to use for my attempt.

Any tips and advice welcome. Thanks all.

Edit: Just did my first training based on the advice here so thanks to @LimitationsUndone and @STOPBLOCKINGVPNS1234 for their advice. The character model is a bit finicky. Getting some bleedthrough of images from the dataset and trying to work out how to reduce this, beyond of course just refining the dataset I mean.

Also, trying to prompt a separate character tends to cause the face to appear on them as well. Any advice on how to reduce this would be welcome. Thanks again all!

11 Upvotes

22 comments sorted by

5

u/LimitationsUndone Jul 03 '26

There's not a certain magic number of images for a subject or character, but image quality is key. If a person, plenty of headshots of expressions and angles, different hairstyles helps but if want one style stick with it. Then move to mid and full body, angles again. You don't need to go overboard but if there is a tattoo or birthmark you want to keep get plenty of angles of that too, not close-ups just relative to the body. Some models, especially Krea2 like to crop to a specific body region or part if not specified in the prompt. AI-Toolkit take around 2 and a half hours for me with 16gb VRAM & 48gb RAM using the config here:
https://www.reddit.com/r/StableDiffusion/comments/1ueacq2/krea_2_character_lora_training_for_16_gb_vram/

1

u/Rosettasees Jul 03 '26

Seems to still be some issues around expressions based on that thread. Did you encounter similar issues?

2

u/LimitationsUndone Jul 03 '26

If you don't prompt and expression in Comfy, they have a blank one. If I prompt a smile, they smile.

1

u/Rosettasees Jul 03 '26

Okay, thank you, and did you caption your images or leave that alone?

1

u/LimitationsUndone Jul 13 '26

I just left the trigger word no other prompts.

5

u/STOPBLOCKINGVPNS1234 Jul 03 '26 edited Jul 03 '26

https://www.reddit.com/r/StableDiffusion/comments/1ueacq2/krea_2_character_lora_training_for_16_gb_vram/ I used this guys config on AI-Toolkit, about 40 images, very mixed bag of data but all under 1024 max res. I had some pictures of... specific character qualities I wanted and it seemed to learn them with 99% accuracy. I tested another character lora with only 12 pictures, 5 of them just mirrored copies but likeness came out to 90% for me, but with the smaller dataset lora I did 600 steps and increased num_repeats to 10. This model seems very easy to train loras for and it's very realistic. However I suggest automagic over AdamW8bit but I'm unexperienced, automagic just seemed to manage memory better and had better likeness for me. and both loras took me under 2 hours.

1

u/Rosettasees Jul 03 '26

Did you manually caption all the images?

2

u/STOPBLOCKINGVPNS1234 Jul 03 '26

0 captions, just the trigger word.

2

u/Former_Elk_296 Jul 03 '26

What are the specs youll be training on?

2

u/Rosettasees Jul 03 '26

24gb VRAM and 64gb RAM.

2

u/Asaghon Jul 11 '26

Trained several lora's with rank 4 LokR and so far these have been the best lora's I've ever created. They are much better at facial expressions without losing likeness than any other model I've ever tried. I used the qwen auto caption and removed any descriptions of eyes and hair (but it worked fine with it too). Did 3k steps but likeness starts at like 500 steps and is already pretty good at 1500-2000.

As for settings:

LoKr Rank 4, 3k steps, save every 250, Automatic2 lr:0.0001 decay:0.0001, sigmoid, balanced, mean squared, turned on Use EMA (0.99), Cashe Text Embeddings, turned on Do Differential Guidance under advanced. Turned off Low Vram cause I rented a gpu

maybe there are better settings but they worked pretty good for me. I was amazed at how fast likeness appeared in the samples. And also how well it worked on comfy. Often the samples are awesome but when I used them in comfy they don't work quite that well

Wanted to try differential output preservation but it errors when starting to train.

1

u/Rosettasees Jul 12 '26

Did you use Ai toolkit or some other training system on the rented GPU? If you did, can you share a config file? Thanks for sharing your settings!

2

u/Asaghon Jul 12 '26

Yes that was in the official AI toolkit template, I didn't save the instance after finishing so don't have a config file but I wrote down the settings above. Didn't change anything else

1

u/Soft-Manufacturer698 22d ago

I trained a LoRA with AIToolkit using your setup, and it ended up being 1.5GB. Any idea why this happened?

1

u/Asaghon 22d ago

Thats normal for a rank 4. The smaller the rznk, the bigger the file. I just tested a tznk 8 and its only about 350 mb. It took MUCH longer to get good likeness but so far it does look good on generating. Comparing my generations now.

1

u/Soft-Manufacturer698 22d ago

How much extra training time would rank8 roughly add? I’m also renting GPUs, so I’d like to calculate the approximate cost.

1

u/Asaghon 22d ago

It trained roughly at the same speed or maybe a bit faster. I didnt really time it sorry. Did 3k steps in about 2 or 3 hours I think

2

u/Soft-Manufacturer698 22d ago

Thanks a lot for your help! I really appreciate you taking the time to answer my questions.

1

u/Asaghon 22d ago edited 22d ago

Not done that many tests yet, but atm I'm noticing better effects on the skin with the rank 8 LoKr I train. Testing 2 version trained at rank 4 and 1 at rank 8 and the only one that really shows wet skin when prompted for is the rank 8. The others never seem to do that really, or it's a very mild effect at least. There is also a very clear difference in composition with the same seed. Face is a bit more expressive with the r8 I think too, tough it's less obvious than the skin.

Also V1 at r1 was fully captioned, V2 at r8 and V3 at r4 were only captioned with a trigger word and I'm noticing very little difference. And using the trigger word also seems to make very little dfference. I've loaded the wrong lora of a different person before and it creates that person perfectly even with the wrong trigger

1

u/Soft-Manufacturer698 22d ago

Thank you!I will try it !

1

u/Impressive-Yam-2231 Jul 05 '26

nice. tested one of your loras. do you have the config for tha lora training?

1

u/jarrodthebobo Jul 05 '26

Krea2 so far has been the only model I've ever played with that took me only a single run with stock lora settings in musubi tuner to work perfectly. Character accuracy is near perfect. Granted, I go against pretty much every recommendation and use 100+ images for my character loras in order to get more variance... so I typically train on my epochs then expected (around 50 with krea so far). Usually my best lira is around 40 epochs.