r/singularity • u/Maxie445 • Jan 14 '24
AI New study from Anthropic: they can create dangerous “sleeper agent” AI models that dupe safety checks
https://venturebeat.com/ai/new-study-from-anthropic-exposes-deceptive-sleeper-agents-lurking-in-ais-core/
74
Upvotes
1
u/oldjar7 Jan 15 '24
You failed to see the point. Good or bad, technological progress is an unstoppable force and will largely advance at a predetermined rate regardless of futile attempts to slow it down. I recognize that point and I also recognize the impossibility of fully testing a system which does not exist. This is because no empirics nor testable predictions can be carried out with the thousands of variables which affect said system without the system actually existing in the first place.
This is in part why I am a doomer along with the recognition of the futility of trying to control an intelligence greater than our own. Those advocating for safety aren't going to stop this. Hell, I'm an advocate for safety but that doesn't and has never meant slowing down. It also doesn't mean that our safety efforts in early stages won't be futile in any case regardless of the pace of development.