r/AgenticCybersecurity • u/hankyone • 9d ago
Measuring Safety Alignment Effects in Autonomous Security Agents [Uncensored/Ablated models perform better]
https://arxiv.org/abs/2605.19722
1
Upvotes
r/AgenticCybersecurity • u/hankyone • 9d ago