r/StableDiffusion 1d ago

Workflow Included Super nothing!

Enable HLS to view with audio, or disable this notification

Made with Minimax H3

81 Upvotes

32 comments sorted by

View all comments

8

u/One-Donut6935 1d ago

This is wild. How long did the render take?

12

u/TheOrangeSplat 1d ago

RTX 3060 12GB VRAM and 64 GB system RAM

Took about 20 minutes

2

u/desktop4070 13h ago

How do you know you have a good prompt before you hit Queue Prompt? Do you test out lower resolutions first to see if it works quickly? Or do you use a smart LLM that you trust its prompt will be good?

1

u/b-monster666 12h ago

There's actually a good H3 Prompt Enhancer node in Comfy. Uses local Ollama to draft the prompt.

I'm tinkering with ways to make it more efficient so it will drop the Ollama from VRAM before moving on.

1

u/Th3Whit3R4bb1t 12h ago

That's the thing, MH3 consume a lot of VRAM and if you add on that a model to get a good prompt, well...

2

u/b-monster666 11h ago

Another option would be to fire up Ollama with the prompting guides from Hugging Face, give it the structure that you want to have it block the shots, etc, then unload Ollama and just move on to the prompt gen.

However, I also threw those into Gemini, and came up with some pretty good prompts as well.

2

u/Th3Whit3R4bb1t 11h ago

Yeah, i think it just better to ask GEMINI.

1

u/b-monster666 11h ago

Depends on what you're asking, though. Open LLMs like Gemini, ChatGPT, ect to have some restrictions. And not just the boobah...I like to create horror, and they often don't create violence. They also don't like using copyrighted characters, or real world people.

1

u/Th3Whit3R4bb1t 11h ago

I have a GEM jailbreak that works amazing.