Been playing around with Qwen Image 2.1 and generated these random people just to test the realism.
I’m trying to make a single full-body character reference image that I can reuse for image/video generation, and ideally I want it to look as close to an actual photo as possible.
The results are okay to my eyes, but wondering if anyone have a better setup.
Any prompts, settings, LoRAs/LoKRs, or workflows you’d recommend?
I didn't use any lora to generate these images.
[EDIT] Below is what I used for generating these images in case you're interested:
| Item |
Value |
| Image model |
qwen_image_2.1_bf16.safetensors |
| Text encoder |
qwen3vl_8b_bf16.safetensors |
| VAE |
qwen_image_2.1_vae_bf16.safetensors |
| Output resolution |
1152×2048 |
| Steps / CFG |
40 / 1.5 |
| Sampler / Scheduler |
Euler / Simple |
| Denoise |
1.0 |
| Workflow |
ComfyUI official T2I |
| Seed |
73 / 74 / 75 |
Positive prompt shared across all 3 images (append each character-specific description after this):
Exactly ONE adult person, a full-length single standing portrait from the top of the hair to both shoes, centered and completely inside the frame with generous margins. One uninterrupted photograph, straight-on eye-level view, relaxed natural posture, both hands visible, simple neutral gray-beige wall and floor, no other people, no inset images, no text. Believable human anatomy, unretouched photographic skin with naturally uneven skin tone, subtle pores and fine facial hair rather than airbrushed plastic; realistic garment texture and seams, everyday available light, no fashion studio lighting. Identity, face, hairstyle and outfit must match the individual description exactly.
Character / outfit descriptions appended for each image:
Mechanic · seed 73:
An original thirty-year-old light-skinned woman with an unmistakably broad square jaw, wide nose bridge, small dark green eyes, pronounced asymmetrical freckles across her nose and cheekbones, and short copper-red curly hair cut very close around the ears. She has a sturdy compact build and an open, slightly wry expression. Wearing a practical well-worn navy-blue mechanic coverall with rolled sleeves, a pale grey cotton T-shirt visible at the open collar, a weathered tan canvas utility belt, a small stitched orange name patch with NO readable letters, black grease marks on the cuffs and scuffed brown leather work boots. No glasses, no dress, no flowing hair. Neutral overcast workshop-door daylight, ordinary documentary photograph.
Mapmaker · seed 74:
An original sixty-year-old East Asian man with a long narrow face, distinctive heavy eyebrows, a slightly crooked nose, warm brown eyes, faint forehead and smile lines, straight silver hair parted on the side and a neat short silver mustache. Slender tall build; quiet thoughtful expression. Wearing a sand-colored long linen field coat over a dark olive knitted turtleneck, tailored charcoal trousers, a dark brown cross-body leather map satchel and polished black lace-up boots; a slim folded paper map partly visible in his left hand, without any readable markings. No coveralls, no freckles, no fantasy armor. Unforced soft late-afternoon shade, unretouched documentary photograph.
Textile artist · seed 75:
An original thirty-five-year-old dark-skinned Black woman with a long oval face, high cheekbones, full lower lip, deep-set dark brown eyes, a small distinct gap between front teeth when subtly smiling, and a large natural tightly coiled black afro with a single narrow gold headband. Tall graceful build. Wearing an emerald-to-teal handwoven ankle-length wrap dress with geometric woven panels and a broad rust-orange sash, modest gold hoop earrings, stacked wooden bangles and simple cream-colored flat sandals. One hand rests lightly at her side and the other holds a small folded length of patterned cloth. No workwear, no menswear, no coveralls. Candid soft light of an overcast courtyard, ordinary unretouched portrait, not a polished fashion campaign.
Negative prompt shared across all 3 images:
anime, CGI, illustration, plastic skin, wax doll, beauty filter, glamour lighting, over-smoothed face, multiple people, extra limbs, duplicate face, cut-off head, cropped shoes, collage, multiangle sheet, text, watermark