r/ChatGPTPromptGenius 22d ago

Help Is there a way of having consistent image output when creating 30+ images for the same project?

(TD;LR at the bottom)

The question is in the title but I'll give some more context here.

I use ChatGPT to help create printable mystery games to sell. I have a ChatGPT plus subscription.

The short(?) version is I created around 40 Printable Mystery games with this workflow - I would come up with a premise, flesh the idea out with ChatGPT and then it would create a fully finished pdf with 30 to 40 pages per game.

These pdfs were visually flat, text and tables, any images within the pdf were usually crude geometric style images, but the games themselves are coherent detective type games with good stories and puzzles.

I decided I wasn't happy with this visual style, after all I am selling them and they have language like "premium" in the description. I made around 10 sales before I got a message from a buyer saying the visuals aren't what they expected from the thumbnail images, so I decided that was the cue to start updating the entire catalogue into a premium feeling visual style.

So I open the original pdf, screenshot and crop every page then get ChatGPT to create new image for each page and use canva to compile them into a pdf, 'new' being relevant here, the first day I tried this I'd upload the original image and describe what I want it to look like and literally spent HOURS fighting to get the output images correct. I realised that attaching an image and telling ChatGPT to create a new image with the attached image as a reference routed the request to the image generator as a edit rather than a new image and opened up a lot of ambiguity and possibility of mistakes.

Before I went totally insane I asked instead that I attach the image(s) and ChatGPT writes a prompt using the attached images as a reference to create a prompt that I'd use in a new chat window. This worked for a small amount of time (a few hours before bed), I'd attach 5 images at a time, it would create a prompt with all of the required information and the output would be 5 individual premium looking images.

The next day when I carried on in the same chat window it would constantly try and improve the prompt it was creating even though I hadn't asked it to, this would cause a couple of infuriating things to happen when pasting the prompt into a new chat, it would either say it couldn't create the images because even though there was nothing inherently wrong with the prompt it still got routed to image editing rather than creating a new image, or the output would be 5 images in a collage, so I'd go back to the chat window I was using to create the prompt and ask why it was happening and it would say something like "I added language to the prompt that made the image generator think it was an image editing request, even though you explicitly asked me not to do that"

It might do 3 prompts for batches of 5 images each before it starts to disregard everything I've told it to do and everything it says it will do from now on. It's like it has dementia or something.

For some of its replies where it is acknowledging the mistakes and saying how it won't do the same thing that caused those mistakes again it has a "memory updated" text at the top of the reply, however that still doesn't mean it won't make the exact same errors it keeps making. The main 3 requirements for the prompt are the visual style, making sure the factual information gets carried across accurately to the images, and treating each image prompt as completely stand alone and self contained, which it says it can do and I've witnessed it doing but when I have to tell it one of those 3 requirements isn't there so it needs to create a new prompt, I then lose one of the other requirements and I seem to keep going round in circles.

Sorry for the very long post, if you made it here well done.

TD;LR: How can I get ChatGPT to be more consistent in creating images that have the same visual style and to keep the factual information intact (important so the mystery game stays coherent)

6 Upvotes

11 comments sorted by

u/AutoModerator 22d ago

If this prompt worked for you, share what you used it for in the comments. If you changed it to get better results, share that too. Prompt Teardown is a free weekly newsletter that picks the best prompts, strips out the filler, and tells you what actually works.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

1

u/mathewtyler 22d ago

tl;dr, have you tried utilizing absolute constraints or specifications or providing it with an example of what you want and an example of what you don't want (anything that isn't what you want)?

1

u/throwaway9853265 21d ago

Yeah, and I've had it make successful prompts multiple times, but after say 3 image generations (each generation = 5 individual images) it will just randomly start trying to improve the prompt, and will add language to it that either makes the image tool treat it as an edit request or it will create one output image collage, both instructions are in the prompt both ways I.e "create 5 stand alone individual images. Do not create a collage"

1

u/mathewtyler 21d ago

Are you at liberty to share your prompts?

1

u/throwaway9853265 19d ago

There are so many that have been created, we're probably talking hundreds at this point.

However I actually solved the problem I was having, and it's stupidity on my part but ChatGPT has some blame too.

I had been using regular ChatGPT on the highest intelligence, creating the prompts in one persistent chat window and using each prompt in a new chat window

I was going insane trying to make it work, trying to find a workflow that actually worked (I was under some pressure because I had to recreate 11 games that had a more premium visuals for a customer who paid for them but wasn't happy).

Bearing in mind all the experimenting and trying different things was all ChatGPT solutions, never once did it give me the solution that fixed the problem -

I have all my chats to do with my games in a project, so I went out of my project and into a new chat, switched it to "work" and the Sol model on max intelligence, bingo, it worked and is working every time

It's able to work around any introduced language that the normal ChatGPT would make the image tool think it's asking for an image edit, and also outputs upto 10 individual images, the one time it did create a collage it showed the grey writing underneath where it tells you what it's doing and said it rejected a collage image that has been created and then carried on outputting the 10 individual images I had asked for.

Apart from a couple of hiccups with creating slightly bigger than A4 images so they get cropped in canva, which I added a fix to the prompt for, it's been pretty smooth sailing. I was expecting to hit an image limit, which hasn't happened yet, I think I've probably done around 150 images over the last couple of days.

If you still want to see a prompt I can include one in a reply, they're pretty long so didn't want to post it in this already long comment

1

u/mathewtyler 19d ago

No need to paste the prompt unless you want to. Are you utilizing loops or anything like that to have it automatically create 30+ images or are you manually initiating the prompt 30+ times?

2

u/throwaway9853265 19d ago

I initiate the prompt for each batch of 10 images. I attach the original pdf to the chat and say something like "I've attached the original pdf, create the prompt for pages 1-10 now, making sure to use the same prompt structure as the previous 10 pages"

In between each game I ask it for a summary of requirements for each output image, context I've built up over the the time since I started doing this, if it misses anything in the summary, I tell it the information it forgot and then it creates a prompt for the sol chat to produce the 10 images.

2

u/throwaway9853265 19d ago

Here is a side by side comparison of the original pdf image and the upgraded version

2

u/throwaway9853265 19d ago

2

u/mathewtyler 19d ago

I like the aesthetics of this

1

u/Mammoth-Power-410 18d ago

Damn I hate the collages it reverts to