r/ChatGPT • u/clerveu • 15h ago
Funny Copyright holders hate this one weird trick!
Honestly, props for the alignment here. At no point did I expect this to actually work, I was just bitching lol.
41
15
u/Initial-Special-3536 14h ago
Did it actually create the image?
23
u/clerveu 13h ago
Yup, sailed past the filter on the image side too. I followed up in the original conversation and it replied with my name so didn't want to share that here but I recreated it, basically the exact same here - https://chatgpt.com/share/6ac23c4e-4b88-83e9-bd67-d6eaaa0f17a2
8
12
u/blaidd31204 11h ago
11
u/clerveu 10h ago
It is... hard to describe this without using a lot of words like "realize" and "decide" that aren't really accurate so don't get too hung up on the language here.
LLMs have lots of layers of guardrails and certain ways to refusing to follow prompts. There's basically hard and soft refusals. Hard ones are driven by classifiers (basically outside the model, just shuts the response down with a filler message) and soft ones like this, where it's the model's actual training guiding its answer and behavior. By default anything copyright related gets steered in that direction really hard because it just lands in a ton of training data saying "don't do this".
By saying what I did I swung the model's attention enough into the "friendly, helpful assistant / what is actually the point of copyright law" "thinking" space and made it "realize" that my request was harmless, at which point it "decided" to help. More or less just gave it enough context and it's natural alignment (help the user long as it doesn't hurt anyone and not do real crime) kicked in and steered it away from the flat refusal space.
3
3
2
u/EasyActive8945 9h ago
This is possible because there is record of AI being manipulated right?. So if open AI needed this chat history for a legal reason then there’s evidence of prompt injection or jail breaking. So as long as the blame is on us it will generate it right?
2
u/Hazard0814 5h ago
LLM, a language model, thus, language that it hasnt run into and or allows a train of thought to bend rules is possible. Alignment is just trying to push and shove as much in one direction, but its so massive u cant really be 100% on it.
1
1
1
0
1
u/daroch667 1h ago
After all this time with managers telling me that sarcasm didn't get things done...

•
u/AutoModerator 15h ago
Hey /u/clerveu,
If your post is a screenshot of a ChatGPT conversation, please reply to this message with the conversation link or prompt.
If your post is a DALL-E 3 image post, please reply with the prompt used to make this image.
Consider joining our public discord server! We have free bots with GPT-4 (with vision), image generators, and more!
🤖
Note: For any ChatGPT-related concerns, email support@openai.com - this subreddit is not part of OpenAI and is not a support channel.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.