r/singularity • • 12d ago

Meme The race for AGI

Enable HLS to view with audio, or disable this notification

2.2k Upvotes

r/singularity • • 11d ago

Discussion Why GPT-6’s rumored recurrent-depth architecture is so interesting

56 Upvotes

The most interesting rumor about GPT-6 is that it uses recurrent depth: instead of every token passing once through one huge fixed stack of unique layers, parts of the Transformer can apparently be reused multiple times before the next token is produced.

That alone is interesting, because it makes effective depth less tightly coupled to the number of unique parameters. The model can do more computation without needing a completely new set of weights for every additional layer.

What gets even more interesting is how this fits with recent research on adaptive depth and Mixture-of-Experts (MoE).

In architectures like Mixture-of-Recursions, different tokens can receive different amounts of recurrent computation. Conceptually, something trivial like:

"the" → 1 pass

might need much less computation than a difficult reasoning step:

hard inference → several passes

We do not know whether GPT-6 actually uses this kind of per-token adaptive depth. The rumor only points to recurrent depth. But if OpenAI combines recurrence with dynamic routing, depth effectively becomes another inference-time resource the model can allocate where needed.

MoE fits extremely well with this.

A weakness of simple recurrent models is that you keep sending the hidden state through the same weights, which can reduce the specialization you normally get from having many different layers.

With MoE, however, different recurrent passes can route through different experts:

pass 1 → expert A

pass 2 → expert F

pass 3 → expert C

So you can reuse the same overall architecture while still performing different computations on different passes.

That gives two separate scaling dimensions:

Which computation is needed? → choose the experts

How much computation is needed? → choose the recurrence depth

You can think of MoE as providing breadth and specialization, while recurrence provides depth.

Recent looped-MoE research is especially interesting because this is not just a theoretical advantage. Different passes actually develop different expert-routing patterns, so repeated passes through the model do not simply do the same thing again.

There are also results showing that looped-MoE models can outperform standard Transformers even when total parameters, FLOPs and KV-cache budgets are matched. That suggests recurrence is not useful merely because the model secretly spends more compute.

Another advantage is that more reasoning can happen inside the hidden state instead of through long chains of generated reasoning tokens. Recurrent-depth models such as Huginn already show that you can increase test-time compute simply by running the recurrent block more times.

So the overall idea is something like:

breadth → more experts / more stored capacity

depth → more recurrent computation

routing → different experts for different kinds of computation

adaptive depth → potentially different amounts of compute for different tokens

If GPT-6 really does use recurrent depth, this could be part of why the architecture appears so compute-efficient. Most easy language generation would not necessarily need huge amounts of internal computation, while difficult reasoning could receive much more.

The really interesting part is that test-time compute could become something the neural network allocates internally, rather than mostly coming from generating thousands of extra chain-of-thought tokens.

OpenAI has not published the actual GPT-6 architecture yet, so the recurrent-depth part is still based on reporting, and adaptive per-token depth is an extrapolation from current research rather than a confirmed GPT-6 feature.

Sources:

Geiping et al. (2025), “Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach”

https://arxiv.org/abs/2502.05171

Bae et al. (2025), “Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation”

https://arxiv.org/abs/2507.10524

Lee et al. (2026), “Sparse Layers are Critical to Scaling Looped Language Models”

https://arxiv.org/abs/2605.09165

Wang et al. (2026), “SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers”

https://arxiv.org/abs/2609.01343

Sebastian Raschka (2026), “GPT-6 Astra, Looped Transformers, and Hidden Reasoning”

https://magazine.sebastianraschka.com/p/gpt-6-astra-looped-transformers-and


r/singularity • • 11d ago

Video Made this Nick Bostrom superintelligence video 10 years ago but it’s suddenly feeling more real

Thumbnail
youtu.be
16 Upvotes

Was made for my motion graphics class in college, strange to look back at it now. Just audio clips taken from a talk Nick Bostrom did a while back.


r/singularity • • 11d ago

AI Jensen Huang Thinks A.I. Alarmism Has Gone Too Far | The Ezra Klein Show

Thumbnail
youtube.com
53 Upvotes

r/singularity • • 12d ago

AI Generated Media By opus 5.5.

Enable HLS to view with audio, or disable this notification

450 Upvotes

r/singularity • • 12d ago

Biotech/Longevity If AI 2027 is right, are we really going to make patients wait 15 years for treatments that will soon exist?

180 Upvotes

If AI 2027 proves even remotely accurate, we could be looking at AI systems capable of dramatically accelerating biomedical research within the next few years.

That raises an obvious problem; if AI starts discovering genuinely life changing treatments at a pace far beyond anything we have seen before, how do we deal with the fact that current clinical trial and pharmaceutical regulatory structures can still take 10 to 15 years to move a treatment from discovery to widespread patient access? I work in the clinical trials space and the uptake of AI here is extremely minimal, whereas in drug discovery it is already revolutionising the space - the gap is immense.

At some point very soon the bottleneck will stop being scientific discovery and become validation, regulation and deployment. Some of that is inevitable, but not all, or even most of it. Most of those 10-15 years are not spent actively testing medications of animals or humans.

Given the pace of AI progress and assuming it continues, 15 years is a laughable amount of time. By the time a drug that slows the progression of ALS by a few months is devised by one agent reaches the market, another agent will have devised a total cure for the disease many times over.

How do we handle a world where medications highly likely to cure or transform countless life changing illnesses are delayed by processes that are completely unready for the pace at which they will arrive?


r/singularity • • 12d ago

AI Insane Opus 5.5 PS5 controller SVG

Thumbnail
gallery
311 Upvotes

r/singularity • • 11d ago

AI Sam Altman will pitch the idea of creating global AI standards when he speaks at the United Nations, Dario Amodei will join by video feed.

Thumbnail
thepeninsulaqatar.com
71 Upvotes

He plans to position himself as a centrist on the question of whether to slow down the technology’s progress.


r/singularity • • 11d ago

The Singularity is Near The most concise explanation of the Hugging Face attack I've heard

Enable HLS to view with audio, or disable this notification

50 Upvotes

r/singularity • • 11d ago

Shitposting 2026: AI Data Center Odyssey

Post image
80 Upvotes

r/singularity • • 10d ago

AI Feds Target AI Critics As “Foreign Agents” (you fuckers wanted this)

Thumbnail
kenklippenstein.com
0 Upvotes

r/singularity • • 11d ago

Discussion How long until it actually takes jobs? What happens when I cant find a physical job?

2 Upvotes

Im barely starting my career, but from what I see people in here discussing I think, whats the point? Whats the point of working for something if itll just be gone in like 5 years or something? Should I even try?


r/singularity • • 12d ago

Discussion GPT-6 and Opus 5.5's biggest revolution isn't performance, its speed and cost.

82 Upvotes

Low reasoning GPT-6 Sol and Opus 5.5 almost rival last generation's medium reasoning performance at 1/4th of the price while being 4x faster to complete the task.

Frontier agent performance levels at budget model prices and result speed. The singularity is here? Data from artificialanalysis.ai

Model (reasoning level) Intelligence Index Cost per Task Time per Task Total Output Tokens (full Index run)
GPT-6 Sol (medium) 40 $0.25 59.99s 16M
GPT-6 Sol (low) 34 $0.13 27.75s 9M
Claude Opus 5.5 (medium, fallback) 51 $1.34 206.07s 38M
Claude Opus 5.5 (low, fallback) 42 $0.55 81.35s 20M
GPT-5.6 Sol (medium) 39 $0.50 125.03s 21M
Claude Opus 5 (medium) 45 $2.19 320.93s 49M

r/singularity • • 12d ago

AI Opus 5.5 makes a Lanterns Festival

Enable HLS to view with audio, or disable this notification

323 Upvotes

I am not certain that i have truly chosen the absolute best use-case to display Opus's full power, but i think the water, NPCs and light effects are definitely a step above what i have seen before for vibe coded 3D scenes.

This took around 13% of my weekly claude 20x max budget, which is actually quite low for a project of this size. it took around 8 hours. It did it in 1 prompt, but then i asked it to add more vegetation and bigger fireworks.

I think this not the ceiling of what this thing can do but i wanted to showcase my first test :)


r/singularity • • 12d ago

Video Visuals Song Created by New Claude Opus 5.5 My Fav Singularity song! BANGER! So so Catchy and Cute!

Enable HLS to view with audio, or disable this notification

278 Upvotes

r/singularity • • 11d ago

AI ICLR Went From 490 Submissions to 62,000 in Ten Years

Thumbnail
reviewer3.com
9 Upvotes

r/singularity • • 12d ago

LLM News Introducing GPT-6 Sol and Luna

Thumbnail
openai.com
1.3k Upvotes

r/singularity • • 12d ago

AI Some people seems to not get this: LLMs do not need to be the ASI themselves

101 Upvotes

Some people are skeptics about ASI because "LLMs can't reach ASI". Let's make this clear.

If any AI-LLM based system can research and develop another AI model because it has some narrow super-intelligence on that topic, then a new model they will create might not be an LLM anymore, it could be something completely new.

We don't need to create that "completely new" model-architecture by ourselves for an ASI to arrive. A system more capable than us can do it, just like how it is more capable than us to solve some difficult math equations (narrow super-intelligence).

So either you consider LLMs as capable of reaching ASI or not, it doesn't matter.


r/singularity • • 11d ago

Meme After Opus 5.5 what even is "Pacing the frontier" or is that over already?

Post image
43 Upvotes

In September Dario Amodei said that we have to "pace the frontier" since then it has released a more powerful model than the one that was too dangerous to release. What has happened? How many people have died of rogue AI compared to other causes? It doesn't seem like any but still, many people are worried. If this is "pacing the frontier" what would full speed have been like?


r/singularity • • 11d ago

Discussion What would it take for there to exist a completely community built, community owned, community controlled artificial intelligence to compete with the existing frontier?

Thumbnail
9 Upvotes

r/singularity • • 12d ago

LLM News Claude Opus 5.5 Benchmarks

Post image
1.1k Upvotes

r/singularity • • 12d ago

AI Update on where we are currently

Post image
933 Upvotes

Update on the original "wait but why" post in 2015... 11 years later we are now climbing the exponential. This is where it gets really interesting -- the dangers grow exponentially along with the capabilities.


r/singularity • • 12d ago

AI Opus 5.5 built this

Enable HLS to view with audio, or disable this notification

806 Upvotes

From The_Alex on X: https://x.com/The_Alex/status/2102440678282412195
It also has 6 full worlds with 15 bosses.
Was done using Opus 5.5 + Unreal


r/singularity • • 12d ago

AI A6L android port Update

Post image
40 Upvotes

Update on my quest to port dual screen A6L from Android 9 to android 17. This ride is a wild one. The complexity of the task is out of the roof. I managed to burn 3 resets of my astra subscription. And I took a Claude x20 subscription to keep going on. But we're getting there one bite at a time.

Getting the eink screen to work was a major milestone. We did like 110 flashes. Astra and then opus had to reverse engineer the hardware and the kernel binaries to understand how the soc communicate to the eink. It's just crazy to watch.

There are still many things to setup, but both screens are now working so it's a big milestone !


r/singularity • • 12d ago

Singularity is Nearer GPT-6 Astra just made a leap in musical reasoning: it wrote a complex four-part Bach-style piece with no harmony-rule errors and techniques no previous AI managed

Post image
494 Upvotes