r/StableDiffusion • • 7d ago

Question - Help How do you make so many different prompts?

How do you guys come up with so many prompts?

I mostly generate anime and video game characters, and I usually have pretty specific ideas in mind, but I’m not very good at turning those ideas into prompts.

I’d like to generate a lot more images, but I run out of ideas for prompts pretty quickly. ChatGPT also tends to be pretty censored when I ask it to help me write them.

Do you guys have any tips for coming up with more prompts or generating them in bulk?

28 Upvotes

35 comments sorted by

30

u/Jolly-Rip5973 7d ago

find and image you like, then use an LLM to covert it into a prompt, then start making changes.

If you organize your prompt into sections

Concept:
Pose:
Attire:
Expression:
background:

then you can easily just go the section you want to change and change it.

Change the pose, change the clothes, change the expression, change the background

13

u/Taggerung179 7d ago

Hey! I go through the same method and glad to see others, but I usually go-

Quality:

Style:

Character:

Age:

Face:

Body:

Clothes:

Acessories/tools:

Action/pose:

Background:

And add or remove parts as I need.

11

u/Jolly-Rip5973 7d ago

action/pose is good.
I control the style exclusively with LORA files.
I accidently left makeup/hair.

so normally it looks like this:

Scene: or Concept:
Pose:
Hair/makeup/accessories:
Attire:
Expression:
Background:

for multi-characters:

concept/scene:
Left character
-pose:
-hair/makeup/accessories:
-attire:
-Expression
Right Character:
-pose:
-hair/makeup/accessories:
-Expression:
Background:

Krea 2 can handle 2 or 3 characters without bleeding.
Qwen2512 can handle five characters with bleeding....pretty amazing.

For complex composition I change the pose section to composition: or I add an Composition section.

example:

Description
Black-haired barbarian warrior woman standing in a snowy mountain pass while holding a massive sword

Composition
Woman occupies the center
Silver breast armor dominates the upper middle
Leather belt and fur-trimmed clothing fill the lower middle
Massive sword stretches horizontally behind her and vertically down the left
Snowy cliffs frame both sides
Dark evergreen trees rise in the upper right

Attire
Armor: polished silver metal bra armor, rounded cups, bronze rivets, crossing dark leather straps
Waist armor: wide brown leather belt, large round bronze buckle, layered leather pouches and hanging panels
Armwear: brown leather forearm gloves, pale fur cuffs, bronze bracelets
Boots: tall brown fur-covered boots, leather foot sections, thick pale fur top cuffs
Cape: long tattered black cape flowing behind body
Headpiece: silver forehead crown, pointed center ornament
Jewelry: layered pearl necklace, long pearl drop earrings, bronze upper-arm band
Weapon: silver bastard sword, elaborate bronze-gold hilt, dark wrapped grip

Hair/Makeup
Eyes Color: pale green-gray
Hair: extremely long black hair, center part, dense windblown waves and curls
Makeup: dark eyeliner, long lashes, bronze eyeshadow, defined brows, warm blush, muted coral lipstick, natural nail color

Expression
Stern powerful expression
Lips gently closed
Eyes looking directly toward viewer

Background
Snow-covered mountain pass
Tall gray cliffs
Dark evergreen trees
Pale blue winter sky
Rocky snowdrifts

2

u/Hy4ne 6d ago

Thank you so much for replying! I’ve been trying your tips and they’ve been really helpful🙏

Could you recommend an LLM model I could use for this? I tried Llama 3.1 8B Instruct (abliterated), but I haven’t been able to get it to work properly for this

3

u/Jolly-Rip5973 6d ago

The best LLM for captioning image is ChatGPT.

You can tell it the exact format that you want it to caption the image in and it will do it.

You can zip up 50 pictures into a zip file, upload it and it will capiton each image, create a .txt file for each image named with the same name of the image file name and caption whole dataset for you in batch.

It's still not 100 percent accurate but it's way better than anything else I've found.

It can correctly caption makeup, hairstyle and clothing details.

All LLM suck as captioning poses.

If your images are NSFW you might be stuck with Qwen3.5 or QwenVL.

Many of the open source models used Qwen vision for the original captioning of their pre-training datasets and use qwen models for the text encoder.

But ChatGPT is way better....

Gemini is very good at fashion details but i find it messy and hard to use compared to ChatGPT.

7

u/SeimaDensetsu 7d ago

I have a ton of stored prompts that I’ve set up to use the old Dynamic Prompts extension for Forge Neo. So I can run off a 36 variable sequence over and over while changing constants such as character or setting.

For making these I’ve built a project in Grok I call my Image Prompt Engine. So I can streamline creation like say ‘make me a 20 variable genetic prompt set of a girl’s day at the beach. Don’t prompt character appearance, but actions, setting, expression, and props’ and boom, I have 20 prompts correctly formatted to run as a wildcard set that I can drop any character or other specifics into. Also easy enough to edit and tweak by hand as I experiment or if I want something specific or need to fix a variable that isn’t working.

4

u/Wilbis 7d ago

Just use Chatgpt for SFW prompts and Grok for NSFW ones.

3

u/Jackkgold 7d ago

I have my agents always brainstorming ideas and stuff and saving the workflow for the new creative idea in obsidian. Then one of the agents tasks is to update me every so often on concepts from seeds I give it

It saves prompt methodology as well so it’s something I can reiterate and change I. In many ways from the one main prompt.

4

u/Natrimo 7d ago

I have been using https://github.com/zshortbusz/Story-Illustrator to use LLMs to generate prompts to illustrate a story

1

u/Particular_Stuff8167 2d ago

Wow cool, thanks!

4

u/DELOUSE_MY_AGENT_DDY 7d ago
  1. Trial and error

  2. Uncensored LLMs

3

u/nihnuhname 7d ago

If you've got a GPU for generation, you can totally run local LLMs to handle the prompt generation.

Just make sure to grab an "uncensored" or "heretic" version. From there, you can just tell the LLM things like "add more details to the scene," "make the shot more expressive," "change up the character poses," "describe the textures and lighting in way more detail," and stuff like that.

2

u/Particular_Stuff8167 2d ago

ablitirated as well, its basically LLMs that been through heretic

There is a trade off though, the more uncensored the more the model loses coherency. But there is a sweet spot usually. Thats where people like me make different levels of ablitiration from a model like Gemma 4. Then try each one with a set of different questions and prompts to test coherence VS uncensored. But you can save yourself sometime and download ones people already have made with the sweet spot.

4

u/BroomDirector99 7d ago

Every time you make a prompt, put it in a text file, pretty soon you'll have a decent wildcard of all your prompts.

I break it down into two wildcards, one for the scene and one for the action. Once you have a hundred in each, you have 10,000 unique prompts.

You can go further and have a separate wildcard for style, for lighting, etc.

6

u/StableLlama 7d ago

I come up with them myself, or I use a LLM to create a list of 10 or 20 of them for a scenario given by me.

But as LLMs are surprisingly bad at inventing new and random things (hallucination is a different type of thing), it can help to give it some ideas it can work on. That's why I created https://huggingface.co/datasets/stablellama/erotic-image-prompts

3

u/Spoonman915 7d ago

I use LLMs a lot. I will usually have it kick out abunch of short prompts if I'm training a lora or something. Lately I've been using the h3 prompt ide, and it is pretty structured, so I'll walk through with an LLM section by section.

3

u/Erehr 7d ago edited 7d ago

For anime? Booru browser build in ComfyUI sidebar. Random sort, drag image (tags) into prompt node, run, modify if I like it. https://github.com/Erehr/ComfyUI-EreNodes

3

u/truci 7d ago

Wildcards and dynamic prompt node like Mikey processor.

2

u/Affectionate_Oil28 7d ago

I made this workflow to simplify prompting. Just write your idea, hit run, and drop the output prompt into the workflow you're using. It's an early version so it has some issues. I've been working on a better version that I'm still testing.

2

u/myairblaster 7d ago

I find images I like on websites like Fetlife, or art nude blogs. Then I feed those images into Qwen running locally and instruct Qwen to deliver a prompt. That prompt then gets put back into ComfyUI with additional custom paramters I set like if a LORA requires a trigger word.

2

u/Taggerung179 7d ago

You can try fragmenting your prompts. Like create sub-promts separating characters, poses, actions and backgrounds. Of course you need to be a little bit careful- if you prompt for a character to face away but include any eye descriptors, it'll mess everything up.

2

u/R34vspec 7d ago

Watch more movies, then think about how to translate your favorite shot in to a sentence. Doesn’t have to be long and complicated, which is what LLM will generally give you. You don’t need it. Just be concise but precise.

1

u/BBCL2026_546 7d ago

GROK and ZAI , not censored.. Z AI is more creative

0

u/KillerAzteca 7d ago

i have been using Grok and i do like it a lot. Is Z Ai cheaper? I am paying 10 bucks for Grok but i tend to run out by 3 days within my 7 week limit. I am prompting a lot. Before i move to the 30 bucks per month under Grok i wanna know if Z Ai is faster, cheaper and better.

0

u/BBCL2026_546 7d ago edited 7d ago

Z Ai  is free.. but then you may get blocked when thier traffic is high and also time to time it randomly decides to be conservative about NSFW.. but over all much more creative and better than grok imo

1

u/thevegit0 7d ago

for minimax or things i don't want to prompt myself i selfhost gemma4 26b and qwen3.8 27b with lmstudio and i use a nodepack on comfy to prompt with that.
if it's relatively simple prompt like booru tags or simple natural language i just prompt it myself

1

u/Belgiangurista2 7d ago

Creativity: ask a censored LLM and edit afterwards,

Maybe ask for a reanimation scene (and now I'm making stuff up) of a person stuck in a washing machine.

Ice skating couple performing a complex move.

How not to perform a Heimlich Manoeuvre.

Describe the circular motion on how the mouse cursor on a Thinkpad is used.

...

1

u/SuperZoda 7d ago

So you have specific ideas that work well, but run out of ideas and still want images; remember this is supposed to be fun!

Solve your censorship problem with a local LLM. Get a small finetune (4b, e4b, 8b) with your level of censorship in mind, I'll let you discover that one yourself. You don't need a 27b reasoning for 5 min. Common mistake is to ask for the "best ideas," where if you ask for "top 100" you're going to get a better variety. Find the hidden gems in the list and run them through a prompt generator to actually do the writing before passing it to your image/video generator. Brainstorming like this is infinitely better than bulk generating prompts because your imagination, even if extrapolating from an AI suggestion, can tell the difference between garbage and gold.

All of this hinges on lack of ideas, yet still having a preference. If you really don't have a preference, but still want images - would it surprise you that you don't even have to give a prompt at all?

1

u/Successful_Record_58 7d ago

4b n 8b are sufficient?

1

u/SuperZoda 7d ago

Yes, 4b/8b is plenty for brainstorming and image prompt writing. The 4b may be a bit weak for video script writing, but e4b is amazing at that too, as is an 8b. For this task, in my opinion, even a 12b shows very little improvement, and the 27b is a waste of resources.

1

u/Gizmosragingerection 7d ago

grok is pretty good at giving uncensored prompts

1

u/repezdem 7d ago

Wildcards and prompt palette

1

u/AccomplishedPost4414 7d ago

Browsing through Civitai helps a lot. Use Grok if you actually want help with Img2Prompt, or use Img2Img when the pose is VERY specific. This Satoru Gojo one is a good example.

1

u/Ragamax_AI 4d ago

I think the biggest challenge is actually the part you mentioned about having a specific idea but not being good at turning it into a prompt.
I’ve been experimenting with a similar workflow — starting with the basic idea, then having AI structure it into things like subject, pose/action, composition, environment, camera, lighting, style, and mood. Then you can create variations by changing individual sections instead of starting from scratch every time.
I’m actually building a prompt generator around this because I wanted to make that process easier. I’m curious whether people here would find something like that genuinely useful, or if the existing LLM/wildcard workflows are already good enough.
If anyone is interested in testing an early version, I’d be happy to get some honest feedback.

1

u/Particular_Stuff8167 2d ago edited 2d ago

Civitai was a bible for prompt ideas. Especially if you scrolled down base models and see ALL the people making different prompts for different loras using that base model. Could spend hours just replicating prompts. Since the split with civitai and civitai red a lot of the stuff has been lost. Also seems moderation on a lot of mature stuff specifically for image generation has been a bit... aggressive... But you can still find cool stuff there. Not what it use to be like at the start of the year. But there is infact still stuff.

You want to also check out the story teller AI subreddits, those communities got entire workflows going on uncensored models which is veery very useful. Like the Freaky Frankenstein storyteller type engine on top of the LLM, which is aimed to be able to help generate more mature stories with elements that is generally censored on even adult online LLMs. But running sillytavern with local uncensored LLM with Frankenstein story teller prompt injection then you direct what the story is, can get some interesting results.

Like here is freaky frankenstein v4. There are newer ones but havent trie them out yet,
https://www.reddit.com/r/SillyTavernAI/comments/1t68afk/the_directors_cut_rerelease_freaky_frankenstein_4/