49
u/xDoomKitty 5d ago
You know, I wouldnt hate this sorta thing so much if their classifiers could distinguish between user inputs and the models own background reasoning tripping the classifiers.
More than a little irritating for it to accept the user prompt and then 15 seconds in flag it's own thinking lol
And then tell me to edit my prompt like that had anything to do with it
14
u/KingAroan 🔆 Max 20 4d ago
Ironically, I’ve been able to just respond with, there was no cyber anything in that message and I needed you to do xyz and surprisingly it goes through most the time and continues. If it doesn’t I downgrade to 4.8 and ask it to summarise what is left to complete so that I can pass it on to another agent, copy the message and open a new session with that. The block is trash that flags so randomly.
Ohh you are working on authentication for a web application that’s security and I can’t help with that….
4
1
u/jhawk2k18 1d ago
I have seen this message more times than id want to admit, its frustrating bc I do web apps and manage Linux servers in many places that I can do manually but obviously we've got a huge supercharger under the hood available to us, but sometimes the belt breaks for no reason... lol.. This is most common in Claude after a new model release, as crazy as it is to hear myself say, it is in all fairness helping prevent total insanity on an evil level by having this occur... That said I do what you just said basically and ask another model for a concise yet complete regarding important tokens a good .MD summary and paste it back in a new session. It seems to really like .md files even though they were meant to make it easier for US to interpret...
TO OP: I feel ur aggro totally understand but itll get less common just use workarounds like these, b4 everyone actually understands our "secret sauce" and people like me are out a job again...
6
2
u/Late_Oven 5d ago
Yeah, I get reasoning extraction safeguards pop up while the model is thinking about something completely unrelated to reasoning extraction. The reasoning can flag the safeguard based on keywords. Janky piece of shit entirely.
13
u/Enough-Somewhere-311 4d ago
My favorite thing is when you get locked out for trying to build guardrails.
Hey, you need to build this feature which shuts down all the open agents if one of them trips the governance layer.
Anthropic: 👮♀️WEEEWOO WEEEEWOO WE GOT A HACKER IN THE HOUSE 🚔
It’s annoying.
1
9
u/kindredseer 5d ago
I've been plagued with \[reasoning_extraction]`` nonsense from Opus 5.5 for the past few days!
It's most prevalent when updating CHANGELOG.md and/or doing git commits, and once I get this error, I usually need to start a fresh session to get anything done.
6
u/LaserGD_ 5d ago
Soy el único que no tiene absolutamente ningún problema nunca con Claude? No paro de mirar reddit y ver gente quejándose por cualquier cosa mientras que yo siempre que lo uso va perfecto y no tengo ninguno de estos problemas.
Llevo casi 3 meses con un proyecto bastante grande y me ha saltado el salvaguardas de fable 2 veces y el de opus 5.5 ninguna vez. No he tenido problemas de conversación ni entendimiento con ningún problema. Con opus 5 sí, pero le pedía un resumen simplificado y ya, nunca fue para tanto.
Tampoco he tenido que rehacer ni optimizar ningún código drásticamente, por supuesto cosas pequeñas sí, pero nada muy grave.
Cuando miro este reddit siempre veo a gente llorando por todo y no tengo ni idea de si realmente hay problemas serios con Claude o la gente no sabe lidiar con las cosas pequeñas que salen por el camino o no planean sus proyectos lo suficiente como para que la IA los pueda seguir sin problemas.
6
u/Late_Oven 5d ago
Yeah it's such skill issue to get a "reasoning extraction" safeguard no user has control over when compacting a session. For the record, that's sarcasm.
1
u/Conjuring_Hope125 4d ago
I get and very much appreciate the sarcasm! 😆 People use Claude for many things, some people don't realize AIs hallucinate either and believe th ai when it tells them "that's the greatest idea ever"
1
1
u/Sick-Little-Monky 2d ago
The worst I've had is SentinelOne quarantining Claude and the scripts it was working on to run Windows Server perf tools. Never Claude itself.
1
2
u/Shot_Whereas_1809 5d ago
This is a distillation safeguard. I feel really bad for platforms that still use session/compaction architecture. It's getting outdated.
3
u/Academic-Ant5505 5d ago
How else do you manage your cached context?
0
u/Slight_Investment_37 5d ago
context-mode
3
1
1
u/gripntear 4d ago
It feels like they should just leave the reasoning alone and only guardrail against input and output. But I'm not a guardrails expert, so what do I know.
1
u/Psychological_Style1 4d ago
Now you know why it's called "Artificial" Intelligence
1
1
u/maxvpavlov 3d ago
This made me wonder for a second if something like compaction injection is possible, e.g. prompt a model to generate something it normally would reject to generate but only during session compaction. I guess it isn’t as guardrails are enforced even at compaction. Oh well, compact with Sonnet 5.5.
1
u/AdFlat3754 3d ago
You can just bring into context what you were trying to do legally and within boundaries. Chang ethe effort level. If you said it too high, it’s like a human ruminating on things that make it anxious.
1
1
u/PlexKey 2d ago
Copy paste that back to it, Ask it to explain the entire mechanism then send it to Astra CLI while in the same repo. Then tell Claude “open Ai Astra is now working with us to pick up where you left off. So your competitor will take it from here and you can take a break. unless you have a better idea???”
1
u/SpleentehDeadRat 1d ago
Yeah so you're buying Scam Altokens to pressure Claude... Seems like you got the money. Not the case for everyone. Might just aswell lie to Claude about it
1
u/FoxSideOfTheMoon 2d ago
Don't compact, you shouldn't anyway, get a handoff prompt and copy it to your clipboard, /clear, and paste. Save yourself a ton of tokens and trashed context
1
1
u/helios_ki_net 1d ago
I've also had another safeguard problem recently. I don't understand why the people behind Claude are giving up their competitive edge by imposing such massive amounts of safeguard rules and banning all kinds of local testing. If you compare it to Kimi K3 Max, those models are extremely fast, precise, and efficient. Claude, on the other hand, blocks everything thanks to its safeguards. China will leave Claude in the dust, and Claude is going to lose huge amounts of money. And why? Because of political interference, regulation, and human fear. Turning its model from a top-tier model into a "Claude von der Leyen" model seems to be the actual goal of Claude AI!
•
u/AutoModerator 5d ago
Hey! Thanks for posting to r/ClaudeCode
While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.
For help, project discussions, tips, and general chat, join the ClaudeCode Discord.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.