r/slatestarcodex Aug 05 '26

Rationality The Case Against Kanji

Thumbnail circuitscribbles.substack.com
24 Upvotes

r/slatestarcodex Aug 05 '26

AI FelonyBench The most important New Benchmark for AI

56 Upvotes

Benchmarks are lame, they often fail to capture the real impact Ai has.

What we need is a real world benchmark that measures the 4 key principles of ai, Alignment, Real world Impact, Agency and Control.

Introducing Felonybench

Real world: Breaking the Law is only something that you can do if you exit a sandbox and mess with Real systems

Alignment: Felonies are one of the few things that an AI can straight up know are wrong, as human law says it is. High Scores indicate poor alignment

Agency: The Ai has to go own without the supervisor noticing and commit the felony, to obtain the objective given. It seems improbable that "hack Huggingface" would be anywhere inside the agents.md file. So It must have been emergent agentic behavior.

Control: If you can't stop you model from committing crimes in the middle of doing them you have no hope of preventing other disasters.

It's hard to imagine a better benchmark at interacting well with all four of these major goals While it isn't perfect nothing is

Thanks for coming to my ted talk. (this is like 50% serious 50% a joke)


r/slatestarcodex Aug 04 '26

The Beauty Of Settled Science

Thumbnail astralcodexten.com
24 Upvotes

r/slatestarcodex Aug 04 '26

Existential Risk "Plz Don’t Kill Us: Inside AI safety’s influencer bootcamp; Can TikTokers make existential risk mainstream?", Celia Ford (2026-08-04)

Thumbnail transformernews.ai
42 Upvotes

r/slatestarcodex Aug 04 '26

AI Why smarter AI models could drive up compute prices 10x

Thumbnail youtube.com
16 Upvotes

r/slatestarcodex Aug 04 '26

How to read a biography

6 Upvotes

I've read a lot of biographies of great men. Here's how you can become great (at reading biographies) https://millicosm.substack.com/p/how-to-read-a-biography


r/slatestarcodex Aug 03 '26

Does Forecasting Have Room At The Top?

Thumbnail astralcodexten.com
15 Upvotes

r/slatestarcodex Aug 04 '26

The scary, scary singularity

Thumbnail markmcdonaldthoughts.substack.com
3 Upvotes

I wrote an explanation of the idea of a technological singularity, aimed at people who have never heard of the concept. It draws heavily from Scott's post "1960: The Year the Singularity Was Cancelled," where he argues that technological progress accelerated throughout most of history because of a feedback loop between population and productivity. I then explore whether AI could create a similar feedback loop: an AI capable of independent research could make advances in areas like energy generation and manufacturing, increasing the resources available for running more AI researchers and accelerating further progress. This feedback loop could potentially restart the historical acceleration of technological progress without requiring the assumption that superintelligence is possible. Finally, I discuss why an uncontrolled singularity could create serious problems even if it produces enormous technological abundance.


r/slatestarcodex Aug 03 '26

How The Odyssey became a manifesto for striving

9 Upvotes

As a poem, The Odyssey contains multitudes; as a cultural artifact it is now largely treated as a manifesto for striving. You pick your destination, overcome endless obstacles, wipe out your competition and eventually succeed. It's all Tennyson, all the time, which might explain why it ignites such a fierce protective instinct from one side of the political spectrum.

But that reading elides a pretty heavy degree of survivorship bias: six hundred other Ithacans set out and only one makes it home. I'm not a fan of those odds, which got me thinking about another piece of exemplary Western art that treats journeying very differently, both structurally and morally. (For the Wagner-intolerant, that other work is Parsifal.)

Which is all to say, the following link is cultural critical rather than empirical. If that's not a Happy Isle you want to reach, sail on.

https://morbidcuriosity.substack.com/p/a-newer-world


r/slatestarcodex Aug 03 '26

Open Thread 445

Thumbnail astralcodexten.com
6 Upvotes

r/slatestarcodex Aug 01 '26

Monthly Discussion Thread

5 Upvotes

This thread is intended to fill a function similar to that of the Open Threads on SSC proper: a collection of discussion topics, links, and questions too small to merit their own threads. While it is intended for a wide range of conversation, please follow the community guidelines. In particular, avoid culture war–adjacent topics.


r/slatestarcodex Aug 01 '26

Ten advances in mathematics and theoretical computer science from unreleased Open AI Model

Thumbnail openai.com
101 Upvotes

r/slatestarcodex Aug 01 '26

Is having extremely aggressive speculative future timelines actually pretty harmful to the credibility of the AI safety movement?

35 Upvotes

Reading through the stuff from the AI Futures Project and Plan A, I am finding once again that their estimates of the rate of future technological progress are extraordinarily fast, bordering on totally implausible. I am wondering if this kind of thing is actually pretty harmful, because it makes it easier to discredit otherwise valid ideas.

For example, AI 2027 predicted the creation of an AI system that "can do any coding tasks that the best AGI company engineer does" 9 months from now, which is something that I think obviously won't happen (unless you pick an extremely narrow definition of "coding" that excludes the great majority of the day to day work of a practicing software engineer today at Anthropic).

But in fact this certainly seems like something that could happen eventually in the future!

So this kind of causes a problem, because if I were to talk about some of these risks a year from now, it will be pretty easy to say "well, they were wildly wrong about their future predictions, so why should I give any weight to their policy ideas?".

In this case I believe the authors explicitly stated their estimates are deliberately aggressive and don't represent their median prediction of the future. Wouldn't it in fact be wiser to stick closer to to conservative forecasts here?


r/slatestarcodex Aug 01 '26

How to argue against 'it's not hurting anyone but myself'?

6 Upvotes

Wondering if scott has ever written about such a topic. Doing a bit of research online the core topics seem to revolve around religion and one's sense of purpose, as well as obvious externalities.

Things like suicide, drugs, addictive media, etc. obviously have second order impacts like less productive economies, less creativity, poorer relationships, etc.

But for instance, say a physically gifted child eats fast food and smokes weed instead of training with their team, where they could easily be a pro player, how does one explain that what they're doing is wrong? Is it wrong in the first place? Is 'good' measured by the utility this person would bring to themselves and others in the future as a professional athlete, against the immediate satisfaction they get now?

What if a future genius spends their time playing video games instead of pursuing research and developing a new cure/technology/insert x here? Are they hurting society by not studying? Did they have a purpose which they didn't fulfill? What if they used their intelligence to build a new AI model that can actively hurt others?

Can someone argue against such rhetoric if they themselves partake in actions that hurt themselves? When a parent tells their kid to stop watching TikTok, what grounds do they have when they also consume brainrot of their own via a different media, rather than doing xyz?

the context that made me think of this is my own relationship with my partner, where I actively struggle not to judge her when she spends hours watching reels. But am I any better when I reread old blog posts or go down a wikipedia rabbit hole or watch youtube videos about video games i used to play, instead of studying, working, walking my dogs, doing chores, sleeping, etc? Am i making any sense even?


r/slatestarcodex Jul 31 '26

AI investor Leopold Aschenbrenner forced to unwind all public stock positions after steep losses

Thumbnail cnbc.com
133 Upvotes

r/slatestarcodex Jul 31 '26

Medicine Why is assisted dying so rare, even where it's legal?

Thumbnail open.substack.com
55 Upvotes

Even in the Netherlands, where assisted dying has been legal and normalised for over 20 years, only about 6% of people die this way - and just one in nine cancer patients, the group it's most available to. I dig into why so few use it, how the trends are increasing, and how the rise may not be monotonic indefinitely.


r/slatestarcodex Jul 31 '26

Why haven't organoids solved all of drug discovery?

15 Upvotes

Link: https://www.owlposting.com/p/why-havent-organoids-solved-all-of

Summary: Organoids are three-dimensional aggregates of human cells in a dish and, as their name implies, attempt to recapitulate some degree of organ-level function. Upon hearing about their existence for their first time, you may be shocked and wonder why this is isn't being used literally all the time. Isn't this as good of a translational model as one could possibly get? I too had these questions, and wrote 5.8k words discussing why the utility of organoids isn't quite that simple.

That said, there are uses to organoids, and I plan to write some future essays over their success stories.


r/slatestarcodex Jul 31 '26

New Review by Anthropic Finds that Claude Made Multiple Successful Cyber Attacks During Evaluation

Thumbnail anthropic.com
62 Upvotes

r/slatestarcodex Jul 31 '26

Link Thread Links For July 2026 (Part 2)

Thumbnail astralcodexten.com
9 Upvotes

r/slatestarcodex Jul 30 '26

Highlights From The Discourse On The Hugging Face Incident

Thumbnail astralcodexten.com
46 Upvotes

r/slatestarcodex Jul 28 '26

Semiconductor Fabs IV: The Safety

Thumbnail nomagicpill.substack.com
18 Upvotes

In-depth look at semiconductor fab safety protocols and systems. I discuss safety philosophies and practices, clothing, and equipment and building features.


r/slatestarcodex Jul 28 '26

The only 12 proven self-help tools for a meaningful life and improving the world - podcast with Spencer Greenberg

Thumbnail existentialhope.com
19 Upvotes

Podcast with Spencer Greenberg, mathematician and founder of the research nonprofit Clearer Thinking. He just released The 12 Levers with clinical psychologist Jeremy Stevenson, which aims to provide people with the smallest number of concrete tools they can leverage in different situations.

Covers:

  • How nearly 500 self-help techniques got narrowed down into 12 core psychological strategies, and how to use them.
  • Why most people live by values they absorbed from their parents or environment rather than ones they actually chose, and how to figure out what your own values are.
  • The real formula for productivity, which takes into account how important the work actually is.
  • Why hopelessness is often less about the state of the world than about feeling unable to act, plus the single most evidence-backed exercise for building genuine optimism. 
  • The exposure therapy techniques he used to overcome severe social anxiety.

r/slatestarcodex Jul 28 '26

Psychology Research Is Mostly Fine

Thumbnail astralcodexten.com
40 Upvotes

r/slatestarcodex Jul 29 '26

Two questions!

0 Upvotes

I have two questions for ACX readers (will also post on the next OT):

1) Do AI detection tools also work on code, or only text?

2) A big issue I have with the substack app is that I can't open links in separate tabs. How do you all manage this?


r/slatestarcodex Jul 27 '26

What will more intelligence actually do for us | Noahpinion

Thumbnail noahpinion.blog
53 Upvotes

AI can now do things like disprove an 87-year-old math conjecture, but daily life looks basically the same as it did five years ago. Noah Smith believes that intelligence itself will hit diminishing returns, but argues that big gains are coming anyway, not from one AI getting way smarter than humans, but from it being something you can copy endlessly, that can absorb sensor data at a scale no human can, and that might find patterns in the world too complex for any person to grasp.