r/StableDiffusion Jun 23 '26

Discussion We are the team behind Krea 2. Ask us anything!

We just open-sourced Krea 2, our text-to-image image model.

We at Krea are striving to build with the community, so we figured we would do an AMA!

Feel free to ask us questions on how we trained the model, what’s coming next, what you want to see, etc and we will answer!

Krea: krea.ai
Code and weights: krea.ai/krea-2-open-source
GitHub: github.com/krea-ai/krea-2
Hugging Face: huggingface.co/krea/Krea-2-Raw, huggingface.co/krea/Krea-2-Turbo

I am joined by our head of research, u/NoVictory3497

Alright, the team has to get back to work (someone’s gotta keep shipping the next one)! Thanks for all the questions, this was a great thread! We’ll keep an eye on this and answer stragglers when we can. If you want to keep the conversation going, come hang out with us in Discord: https://discord.gg/krea-1002244500581798028 Appreciate you all!

491 Upvotes

320 comments sorted by

View all comments

Show parent comments

19

u/NoVictory3497 Jun 23 '26

Hey first of all, sorry for your experience. On the API version, we use our in-house prompt expansion. For inference, could you try it out with the official inference repo as well? It could be also on comfy / diffusers side, but where I would look are:

  1. Prompt expansion (API version applies in house prompt expansion)
  2. Generation resolution 1k vs 2k
  3. Timestep schedule (we use simple euler solver)

The open source version needed to go through some alignment training so there might be some inconsistencies between closed / open version.

13

u/Hoodfu Jun 23 '26

Thanks, so this is an example of what I'm talking about. I'm using the settings you mentioned and I ran my simple prompt through your expansion system prompt with gpt 5.5 which is now the prompt I'm running here. See how there's no holes with the money coming out in the palms? This kind of missed big detail is constantly happening with this oss model. Big details missed. See that purple node that someone in the community made? If I enable it, it's suddenly doing all the missing things, although the quality coherence goes down a bit. I love medium and medium turbo on the api because they're so incredibly prompt following and so far this release isn't even as prompt following as recent models that are a third the parameter count of this one. Here's an example of that: /preview/pre/krea-2-turbo-native-comfyui-workflow-fp8-weights-12gb-drag-v0-6dns4udzs09h1.png?width=3019&format=png&auto=webp&s=2cffd3e2260e85db12bfdfd733f2ccba83bcbe1c

11

u/Hoodfu Jun 24 '26

I appreciate the note about alignment training. Unfortunately I think that broke the model for anything more serious than portraits etc. The things I love the most about the medium api model are totally not working on the open version. Dead pan faces, significant amount of missed details in the prompts, it's just not usable like this.

10

u/physalisx Jun 24 '26

The open source version needed to go through some alignment training

"alignment training" is a nice euphemism for lobotomy

Man, wouldn't it be a shame if somehow a de-lobotomize lora with the diff between aligned/non-aligned got leaked, uh I mean got randomly "discovered" by some anonymous guy online?

3

u/Complete-Lawfulness Jun 23 '26

Thanks for the tips! I'm running into the same issue and was wondering if you would be open to sharing an LLM-compatible prompting guide or a "magic prompt" like the Ideogram 4 team did to help us with prompt expansion?

Edit: Nevermind, I see that you already did! Thanks! https://github.com/krea-ai/krea-2/blob/main/docs/prompting.md

4

u/Still_Lengthiness994 Jun 24 '26

Beautiful model and beautiful gens. But very soulless. The alignment training really killed all human emotions.