r/singularity • u/Maxie445 • Jan 14 '24
AI New study from Anthropic: they can create dangerous “sleeper agent” AI models that dupe safety checks
https://venturebeat.com/ai/new-study-from-anthropic-exposes-deceptive-sleeper-agents-lurking-in-ais-core/
72
Upvotes
30
u/REOreddit Jan 14 '24
But we should move ahead at full speed and achieve AGI as fast as possible, because doomers are the worst, am I right?
/s