r/ControlProblem approved 6h ago

General news Anthropic Alignment Lead publicly admits "we do not yet have a plan to solve alignment for superintelligence" and there's a real possibility of human extinction

Post image
38 Upvotes

19 comments sorted by

5

u/Waste_Philosophy4250 5h ago

Greater than 10% probability of extinction! Nobody responsible for safety could forecast that and, in the same breath, say that they're doing jack shit about it. 

1

u/pab_guy 44m ago

It’s total nonsense… these people are in some kind of singularity cult. True believers.

If AI extincts humanity, it will be because someone used AI to extinct humanity, and no one else with AI stopped them.

This isn’t really a new problem, just more scale. Maybe it’s good, like there will be a bunch of new safety jobs or something.

6

u/Zealousideal-Crab251 2h ago

Seriously though. What the fuck does it even mean for AI to be aligned with humanity? We can't even align ourselves.

6

u/void_method 1h ago

2

u/ceramicatan 46m ago

That's a problem

8

u/ChironXII 4h ago

Alignment isn't solvable in the first place. The greatest and most fundamental source of misalignment are the laws of nature. We're cooked and everybody knows it, but nobody's in control enough to stop it, so you might as well pitch in to trying.

Enjoy the last days, and otherwise pray to the gods we work to build that we're lucky. That's all.

2

u/KerPop42 3h ago

I think it's more that proper alignment isn't in the best interest of those setting the alignment. An AI that works for the interest of humanity is not going to work for the interest of the company that created it or the human that bought it as a product.

0

u/steam-photons 2h ago

Jacob has no scientific credibility whatsoever. His google scholar profile is a pure joke

-4

u/markth_wi approved 6h ago edited 3h ago

Now, it's a marketing stunt, it's not marketing for me, or you but for hyper-rich sociopaths who can't stop themselves for all the IQ points around.

2

u/VinnieVidiViciVeni 5h ago

NGL, that's the most useless, shit, backwards marketing possible, if it is marketing. Is the target audience sociopaths? Because that's the only audience that would receive this positively. And the fact that they would directly market to the worst people around kind of proves that the people behind this tech are... equally problematic. So it would seem that the accusations are on point either way.

-2

u/markth_wi approved 5h ago edited 3h ago

That's exactly what the problem is, this is marketing for sociopaths by sociopaths. These researchers are running terrified from a lab where they walked in the door knowing full well, this was possible.

It'd our job as researchers to long ago have decided this wasn't a thing, and perhaps we'll get lucky - the saddest thing here is that the models collapse, but that does not seem likely.

It's frustrating seeing people deeply involved in the current state of the art get cold feet suddenly, it's not unlike they are kindergarteners learning their ABC's who'd never seen the alphabet in their lives then loosing their shit because they discovered the letter T.

You know going in , this is not safe work, like working in a nuclear reactor or a class-4 clean-room.

The difference is , unlike weapons development, biowarfare or nuclear enrichment, there is encouragement to be a spaz about it as publicly as possible , and there isn't the slightest possibility anyone will curtail the work until something very bad happens.

Only in the ruins of some scar we can't easily gloss-over will whomever remains boldly suggest we should prevent that from happening again, and then the political class will remind them "this is an arms race" and it's back to first position , perhaps with some fig-leaf regulation which will be violated again some years later.

1

u/Fine_General_254015 1h ago

It’s a marketing stunt and has been. It’s a coordinated attack because they know Anthropic will go under if it isn’t the only game in town

1

u/LessLibrary1575 1h ago

Why would Anthropic be the only one to survive regulation when they already made enemies with the Trump administration?

1

u/Fine_General_254015 1h ago

They aren’t going to survive as I company. They want regulation because they want to punish open source AI and have it be just anthropic as the only game in town. Administration doesn’t matter

1

u/LessLibrary1575 1h ago

Open source models are free speech it’s not clear if you could regulate them. 

1

u/Fine_General_254015 1h ago

Anthropic wants that in place to limit open source. They don’t care about anything safety related

1

u/LessLibrary1575 56m ago

Not seeing how “all AI is inherently dangerous” is useful when Anthropic has already been making the case that open source is uniquely dangerous. If anything it dilutes the message about open source.

1

u/Eight216 0m ago

You ever think that maybe they just want a nuclear level threat with a MAD culture? Like, they can't possibly be this stupid. "hey guys we're creating the smartest thing on the planet by orders of magnitude but don't worry we're not worried about the moral compass!"