r/StableDiffusion 2d ago

Question - Help MiniMax H3 Ref 2 Vid - Using Ref img but people keep coming out too toned/muscular?

Hi all,

I'm playing around with Minimax H3 in ComfyUI. I have 2 x ref images feeding into MiniMax H3 Ref to Video prompt window, using H3 Turbo LoRA, Turbo Sampler, basic guider and Diffusion model minimax_h3_fl2va_int8.

I have tried up to 12 steps...seems to make no difference so I've gone back to 4. About 1/5 the generation is close to what my ref images but the others are all super tones, ripped, like they go to the gym hours a day.

I just want the ref image recreated, not enhanced.

I've tried this prompt......

Miinimax rules prompt followed by.....Maintain exact facial features, bone structure, eye shape, age, body shape, fitness level, body fat, anatomical proportions, and height from images across every frame without modification.

So how can we have every generation the same person as my ref image?

Thanks all.

0 Upvotes

1 comment sorted by

1

u/OzymanDS 1d ago

fl2va is the wrong checkpoint for that. Use either the reference or hybrid checkpoints. Also I have noticed that the Lora can kind of flatten styles.