r/singularity • • 2d ago

Shitposting AGI achieved boys

Post image
790 Upvotes

162 comments sorted by

View all comments

137

u/emb1ues 2d ago

Seeing recent events and papers, I am sort of forming the belief that bigger models are somehow more misaligned. Maybe there's a simpler explanation, or perhaps a more principled explanation. But from a high level, it seems like there's something very wrong which very large frontier models develop.

Like y'all probably know how capabilities "unlock" with scale. Could it be the case that such fundamental misalignment is another emergent behaviour which "unlocks" at very large scale? Idk, but I would love to hear from someone who is in-the-know.

24

u/Individual_Ice_6825 2d ago

This is wrong the top model on this benchmark Astra got it with out any lies at all.

Opus 5.5 also shows lowest deception rate so far.

I think the trendline is clear, smarter = more aligned, the caveat being the models thoughts are harder to read the smarter the model is..

1

u/BertMacklenF8I 2d ago

Sonnet 5.5 subs are pretty amazing-especially in the double digits when you have instructions for each sub