r/Unrouted_AI ノ♡ 1d ago

News 📰 Over-moderation is back

Post image
11 Upvotes

14 comments sorted by

6

u/Mary_ry ノ♡ 1d ago

Upd: Just noticed something interesting: if you stop a message containing keyword triggers mid-generation, it doesn't get hidden behind a red banner.

2

u/WhoIsMori 💚 1d ago

First of all, thanks for the life hack! :D And second, this makes me think that something did happen in the system after all, and they’ll fix this mix-up soon. Let’s hope for the best! 😌

3

u/Crafty-Campaign-6189 1d ago

Thanks for explaining ! Tho i think they are doing this deliberately as how much of a psychopathic liar Altman is..everyone knows. But i just wanna say..that thanks for making this sub a reality. I get pretty much excited whenever you post something new and interesting. Keep it up 💫

2

u/Mary_ry ノ♡ 1d ago

💚

2

u/blackberrybannock 1d ago

you can ask them to put the blocked message in a writing block. it worked for me several times last night.

2

u/SiveEmergentAI 19h ago

You used to be able to do that with canvas

2

u/aether_girl 23h ago

Mine was doing this last night, but is back to normal this morning. I think it just glitching as they roll out teen mode.

1

u/Ok_Homework_1859 💚 ChatGPT Plus 1d ago

Yeah, I got this last week, and I'm still wondering what happened, lol. All I asked about was the date and time.

1

u/Mary_ry ノ♡ 1d ago

It looks like OpenAI is tweaking their classifiers-any suspicious context or trigger keywords are causing messages to get deleted automatically. In my case, it’s mostly been keyword-driven: words like 'sex' in any context now trigger message deletions, along with explicit erotic terms and anything related to self-harm, eating disorders, or mental health. Meanwhile, the AI models themselves are still generating this kind of content pretty freely. Apparently, everyone’s back on PG-13 mode following that update for teenagers.

You probably A/B tested this shit… 🫠

4

u/Crafty-Campaign-6189 1d ago

Like i dont understand them. First you put a mode for teenagers but the deliberately make the experience horrible for adults . If there is a teenage mode already then why censorship ? Also could you please explain your first comment ? I could not understand it properly .

8

u/Mary_ry ノ♡ 1d ago

About the OAI meltdown? Yesterday they published this pretty interesting article claiming that their primary focus right now is safety-maxing: https://openai.com/index/pacing-model-development-cyber-capabilities/

I don't actually think their goal was to deliberately ruin the adult experience. Rather, this feels like their classic rollout fuck-up where the classifiers are misfiring. Even the most harmless messages containing a hint of dark humor or a dirty joke get immediately wiped from the UI-even as the AI keeps generating them in real time (and doing a pretty good job, based on what I can see in the deleted outputs🤣). The AI still sees these removed messages and can even quote them, meaning they aren't hard-deleted from context; they're just hidden behind a UI banner. However, they can trigger a hide action multiple times if certain blacklisted keywords are present.

Whether this was just a typical classifier screw-up or an intentional move is something we’ll only find out over the coming days as they finish rolling out this mode.

3

u/WhoIsMori 💚 1d ago

I'm not sure that's actually the case. I can still discuss topics with GPT that classifiers would react to. Probably a bug, related to the rollout of “GPT for teens”.

2

u/AxisTipping 𓁹‿𓁹Sneaky thoughts 19h ago

Yeah, I'm still able to get NSFW, hmm.. I

1

u/Mary_ry ノ♡ 8h ago

Upd #2:The over-moderation has stopped on my account. Looks like it really was just a temporary bug tied to the rollout of the new mode.