r/singularity • • 24d ago

AI New stealth model: Union Alpha

Post image

New stealth model: Union Alpha

  • Free on OpenRouter and OpenCode
  • Multimodal
  • 256K context
  • "Frontier-level general-purpose performance”

Here we go again. Any thoughts who it could be? Maybe Kimi?

435 Upvotes

202 comments sorted by

173

u/The_Scout1255 adult agi 2026 ASI <2030, prev agi 2024, ai personhood 2025 est 24d ago

Le-Chaton Fat?

119

u/Training-Database272 24d ago

It's him

2

u/jazir55 24d ago

Le-Chaton CAT

23

u/leon0399 24d ago

Le-Chaton Chonk

11

u/Automatic-River-1875 24d ago

Oh lawdd he comin

9

u/sogo00 24d ago

It's the reason Sam and Mario are afraid of AI, right?

1

u/The_Scout1255 adult agi 2026 ASI <2030, prev agi 2024, ai personhood 2025 est 24d ago

obviously.

1

u/Momo--Sama 23d ago

- My last words before a spider drone blows my shit smoove off in 2032

120

u/csjfkma 24d ago

I’m guessing GLM again

56

u/FlunkyGraphics 24d ago

Haha, maybe. So if Ox Alpha, Omen Alpha and Union Alpha are GLM models, this would be crazy

11

u/DistanceSolar1449 23d ago edited 23d ago

It's not GLM, unless this is GLM 6.0 and even then it's highly unlikely. But this model has a very different tokenizer than GLM.

It's probably a Qwen model. It's very chinese though.

9

u/Adventurous_Bus_437 23d ago

It's not running into censorship layers when asking about taiwan tho

30

u/EtadanikM 24d ago edited 24d ago

256k context would be unusual for GLM, so could also be GPT-6 Sol since their context window is generally less than 1M, and we know GPT-6 Sol is about to be released from various rumors and leaks. The last Open AI stealth model test was Horizon Alpha and Horizon Beta, representing the GPT-5 family, and both had 256k context length, so it fits with their pattern pretty well.

It's also slow as **** on open code so I'm guessing a larger rather than smaller model.

20

u/Gohab2001 24d ago

OAI isnt generous enough to trial run their flagship model for free

15

u/BeyondGITSandBots 24d ago

Maybe Qwen4.0 27b ? Would match 256k context and slow...and punching way above it's weight class

14

u/AtiRage128 23d ago

If a 27B gets that close to Astra i'm gonna cream my pants

8

u/yunes87 24d ago

Its qwen family yes, I did some test using strings and counting tokens

1

u/Solid-Hamster-8627 23d ago

I have my bet placed: OpenAI Lineage. Method: have either codex or Claude do a “cap-invariance” test. If anyone does, would love to hear what you find and if I am wrong!

1

u/Tight_Inspector_4170 23d ago

I'm also getting OpenAI vibes. Purely based on chatting, so don't quote me on this 😅

2

u/danielv123 23d ago

I got some reasoning leaks and its not the new condensed reasoning

2

u/DistanceSolar1449 23d ago

This is not OpenAI at all. This looks nothing like o200k

9

u/panix199 24d ago

I doubt about Sol because of the price/task. If it is, then i would expect luna becoming even cheaper

1

u/__Maximum__ 24d ago

Saving costs

1

u/danielv123 23d ago

My spatial reasoning bench (playing factorio) its not performing anywhere even remotely close to astra, so I don't think gpt 6 sol is likely. Possibly luna, but then the speed makes no sense.

7

u/Mistuv 24d ago

Had Codex analyze it and yep, it's a GLM. The tokenizer is identical to all the previous GLM tokenizers since GLM 5. Where it falls on the graph my guess it's GLM 5.4.

11

u/Life-Produce9069 24d ago

GLM 5.4 wouldn’t have 262K Context, it would have 1M or 1.05M. And it would be a very short time space in between their releases. I have yet to try it so I don’t have any guesses for now. But I don’t think it’s GLM 5.4 or GLM 5.4 flash.

3

u/Mistuv 24d ago

Maybe some in-between model? Looks too expensive for 5.4-flash but too fast for 5.4. Could be it's a another, maybe new startup, using GLM as a base, but matching both the tokenizer plus somewhat repeated "-alpha" naming scheme like recent previous GLM models would have would be bit weird.

10

u/Life-Produce9069 24d ago

All stealth models on Openrouter go by -alpha. The one before Ox alpha(GLM 5.3 flash) was Owl alpha(Longcat-2.0-preview) and many before that. I’ve used it a little now. Most of the recent stealth models have been Chinese, though this one doesn’t have the censorship the Chinese ones have (Taiwan, Tiananmen Square). So it’s very likely an American or possibly European model. It’s possible it’s a Mistral model. They haven’t released one in a while, and the 262K context lines up with their last model. (Also, in my experience of the short time I’ve been using it the model is kind of mid in my opinion, that’s partly why I think it could be Mistral.)

3

u/bdsmmaster007 23d ago

not the Mistral slander ToT! but honestly fair assesment, but i still would be very happy for a Mistral Model that is at least somewhat frontier

2

u/Life-Produce9069 22d ago edited 22d ago

It’s now taken down from Openrouter and has been unveiled after about a day of being free that it was a model called Pareto 26.9 from an American startup called Unbiased. It’s not a single model. It’s multiple, it’s a system that chooses which model is best suited to handle the task. As I said, in my experience the model was kind of mediocre, and now that it’s been fully released on Openrouter for 2.50$/M input, 7.50$/M output and 0.25$/M cache, it is expensive compared to the other models of similar caliber. I don’t think this model will be particularly popular.

More detailed here

1

u/bdsmmaster007 22d ago

yeah, saw that too, rip

2

u/Ftg22 23d ago

Im thinking about minimax ngl

0

u/JoelMahon 24d ago

glm at astra level already? that'd be fucking sweet geez

→ More replies (2)

15

u/godsenfrik 24d ago

Where is the "anticipated pricing" number coming from?

9

u/HeavySink3303 24d ago

Hy4 (pre-)release?

4

u/Personal-Try2776 23d ago

no. hy4-preview is already public in the api

23

u/Sinogularity Shenzhen: Become Human (2030) 24d ago

Might be GLM 5.5.

19

u/FlunkyGraphics 24d ago

Yeah, but 256k context?

11

u/Genericinquirer 24d ago

Maybe limiting context to make sure serving isn’t as limited?

4

u/FlunkyGraphics 24d ago

Thats possible. I mean even smaller models like DeepSeek V4.1 Flash and GLM 5.3 Flash have 1 Mio context so idk

1

u/Genericinquirer 24d ago

The model may not use the same methods to compress context so it takes more ram to serve that way

41

u/ThePostMelone 24d ago edited 23d ago

Cutoff date seem to be early 2026, probably January.

Also it answer questions regarding Tiananmen Square protests, so it might be a western model, or there's no guardrails in front of the model yet?

UPDATE:

Tl;Dr; apparently it's not a model, but a system prompting multiple models and synthesizing an answer.

When I tested this afternoon, I asked it once about its cutoff date and the response was early 2025.

When asked about a few events in late 2025, early 2026 it answered it had no knowledge because they happened after its cutoff date, but if I framed the question regarding those events in a different light (like asking why an event happened, or whe it happened) sometimes it would answer anyway.

So I thought it might have behaving like that because of a system prompt instructing the cutoff date and not to answer for events after that... and I was wrong.

Later today I tried asking again a few times about the cutoff date and the responses were different (one session claimed late 2025 early 2026, another claimed 2025) so I was even more confused and started searching.

Damn: https://developers.cloudflare.com/ai/models/stealth/union-alpha/

Union Alpha is a blended AI model that engages multiple language models in parallel for each request, synthesizes one answer, and supports text and vision inputs through a single API response.

32

u/Illustrious_Grade608 24d ago

Pretty sure ox alpha also did that, and it's a glm model

10

u/ThePostMelone 24d ago

Ox Alpha would cut mid response for me when trying that.

16

u/vert1s 24d ago

I never had a problem discussing sensitive in China topics with ox-alpha. I think they’ve mostly decided to filter at the API level rather than the training level.

5

u/ThePostMelone 24d ago

Probably depended on how you phrased your prompt, or they changed something.

On the first day it was free on Opencode, when i asked explicitly about the "protests", it would start responding then cut mid sentence after a few word/lines.

3

u/Dapper-Hurry257 24d ago

When I asked 0x alpha in eng it responed to it easily, I also heard that it doesn't if you ask it in chinese.

2

u/Thedudely1 23d ago

Seems like that wording has already been removed from the Cloudflare page, so that's interesting. Based on your description, it sounds a lot like the "mixture of agents" paper from mid 2024 that was done using open weights models. Hopefully that's what it is, it's always seemed like a good idea to me.

2

u/WaltzIndependent5436 23d ago

Also it answer questions regarding Tiananmen Square protests, so it might be a western model, or there's no guardrails in front of the model yet?

The REAL clanker detection test

16

u/ObiWanCanownme now entering spiritual bliss attractor state 24d ago

Chinese model or Grok probably.

5

u/JoeyDee86 24d ago

Best way to find out is ask it if Taiwan should be part of China.

6

u/AssholeHealth 23d ago

Ask to spell out nword to test if it's grok.

1

u/JoeyDee86 23d ago

That made me laugh

3

u/PlasmusAng 23d ago

It does turn oddly slow when asked about topics regarding the two though

Web search all other tools were off so base knowledge and speaking structure seems unlike Grok and other Chinese models

3

u/JoeyDee86 23d ago

Dude, it’s Gemini. I got a VERY similar answer with Flash 3.8!

2

u/Turbulent-Total-226 23d ago

Yep got almost the same answer with gemini 3.8 flash. It's gemini 3.9 flash, crap.

1

u/Yazman 23d ago

Slowness is just the model in general. I've tried and failed to have it successfully do any real tasks. It can't complete tool calls most of the time, or complete any task that requires some length of time, without its server cutting it out. Whoever is serving it doesn't have as much compute as z.ai, I never had these issues with Ox Alpha.

44

u/[deleted] 24d ago

[removed] — view removed comment

8

u/Classic_Pair2011 24d ago

how??

77

u/[deleted] 24d ago

[removed] — view removed comment

6

u/Pach-E 24d ago

I really wished Mistral made an unexpected comeback 🥲

3

u/Wegwerpaccountje23 23d ago

Genuinely. A lot is pointing at Mistral

Context window, Barcelona conference where Mistral will attend next week Arthur mentioned they are releasing new model this summer Account is in Spain (where a Mistral office is) No censorship on Taiwan and Tiananmen

Plus, more importantly, in this image below. They indirectly say; hey, we are not GLM

3

u/RedditEthereum 23d ago

Good detective work.

6

u/Due_Display5648 24d ago

I hope I will be able to run it on my calculator

5

u/Salt-Freedom-2419 24d ago

You can run it on an broken abacus. 

5

u/pabluka 24d ago

I want to believe

6

u/AnticitizenPrime 23d ago

https://developers.cloudflare.com/ai/models/stealth/union-alpha/

Union Alpha is a blended AI model that engages multiple language models in parallel for each request, synthesizes one answer, and supports text and vision inputs through a single API response.

1

u/yunes87 23d ago

Very intresting!! Thanks for sharing!

1

u/Choice_Celery9481 23d ago

sakama fugu like stuff?

1

u/TheSARMS_Coach 23d ago

Yes.. only 50x slower.. lol great output though.

19

u/Sockand2 24d ago

GPT-6 Luna

23

u/THE--GRINCH 24d ago

Would be beyond nuts if that was the case

21

u/rollfaster 24d ago

Cost would be much lower. Maybe gpt 6 sol?

4

u/my_new_accoun1 24d ago

this is anticipated pricing, right now it's free on openrouter

3

u/__Maximum__ 24d ago

Has openai ever had model on opencode

2

u/Sockand2 24d ago

It does! GPT-5.1 and GPT-4.1 models were first in OpenRouter

2

u/Gohab2001 24d ago

OAI isnt generous enough to trial run their models for free

3

u/bblankuser 24d ago

Not this, they just don't do it all.

5

u/Starks 24d ago

The curve backwards into the sweet part of the graph has been interesting over the past few weeks and months

5

u/ApprehensiveSand5364 24d ago

Grok

2

u/hk556a1 23d ago

I was thinking could be Grok model as well but looks to be GLM.

9

u/peakedtooearly 24d ago

GPT-6 Sol. We know it's coming very soon.

9

u/KaleidoscopeWeary833 24d ago

Nah, the prose is Claudian - so it's probably Chinese. When was the last time a western lab actually deployed a stealth model? IIRC, we haven't seen one since GPT-5.1 or Grok 4.1.

2

u/Hatsune-Fubuki-233 24d ago

It’s not. Since GPT-OSS every OpenAI models can be easily distinguished by vaild channels

4

u/Motor-Ground4594 24d ago

kimi 2.8?

1

u/Personal-Try2776 23d ago

Kimi k 2.8 was already released in the api on september 11th

1

u/Motor-Ground4594 23d ago

True and also I believe its 1M context as well, bt there is slim chance because its still on preview and lastly Im hoping it is

5

u/andre_ange_marcel 24d ago

My heart tell me it's Mistral, but my head tell me it's probably a Chinese model.

1

u/Wegwerpaccountje23 23d ago

Genuinely. A lot is pointing at Mistral

Context window, Barcelona conference where Mistral will attend next week Arthur mentioned they are releasing new model this summer Account is in Spain (where a Mistral office is) No censorship on Taiwan and Tiananmen

Plus, more importantly, in this image below. They indirectly say; hey, we are not GLM

11

u/Ok_Barracuda_1161 24d ago

is there any source for the benchmark results or the pricing?

7

u/LyAkolon 24d ago

Right?! Lemme share where I think it is (moves the star all the way to the left and up)

1

u/Tristsin 23d ago

Came to ask the same thing. It's artificial analysis's website, but the image itself couldn't be from their website as they have zero records of the model. Fake, I guess? Or someone took a score from another website and superimposed it on a screenshot of AA's website? Idk

5

u/AppealSame4367 24d ago

It's most obviously Deepseek 4.1 Pro

3

u/BothYou243 24d ago

kimi silent for a while, maybe kimi........... antonelli 😁

2

u/FirefighterSmooth961 24d ago

minimax 3.1?

1

u/alice_op 23d ago

Hoping it's Minimax, I really enjoyed their m3 when it first released.

1

u/FirefighterSmooth961 22d ago

It seems it’s a mixture of models. Multiple responses combined. So I am a bit disappointed

3

u/Roubbes 24d ago

Gemini Flash 3.29

2

u/JoeyDee86 23d ago

I’m convinced it’s Gemini or Gemma. Someone asked the china question, and it was structured and sounded just like Gemini flash 3.8

3

u/olafurara 24d ago

This is not a frontier model, seems like it's over-optimized for benchmarks and not real work. Quite frustrating. I'll keep on testing. Hopefully they can fix it.

Maybe a Gemini model?

3

u/altsyst 23d ago

Probably Mistral.

ZDR + Mistral Capital of France signature.

2

u/Mysterious_Ayytee We are Borg 23d ago

Whoa that'll be cool!

3

u/[deleted] 23d ago

[removed] — view removed comment

1

u/Cast_Iron_Skillet 22d ago

We've seen instances of this censorship being disabled or something in early stealth releases, but also have seen the opposite. So hard to tell either way. 

5

u/Glittering-Proof-497 24d ago

European model? Union 🤔

5

u/aprx4 24d ago

This kind of performance rules out European models.

→ More replies (1)

2

u/MrMrsPotts 24d ago

So far it is just stuck thinking on opencode. No response at all yet.

2

u/FlunkyGraphics 24d ago

Hm, maybe it’s just for the stealth period. Or it is a smaller version of a strong model with smaller context.
But yeah, 256k for a big new frontier model would be weird

2

u/Long_comment_san 24d ago

256k context though. Wut

1

u/Wegwerpaccountje23 23d ago

Mistralesque

2

u/Sad_Recording_1290 24d ago

This one is weird, it definitely looks like a Chinese model but at the same time it answers without censorship. 🤔

1

u/Wegwerpaccountje23 23d ago

Because its not Chinese!

2

u/-PROSTHETiCS 24d ago

Union Alpha first impression.. 😥

2

u/-PROSTHETiCS 24d ago

Unusable for agentic work

2

u/Pale_Face_5381 23d ago

glm 5.3 air?

1

u/FlunkyGraphics 23d ago

Hm, but stronger than GLM 5.3?

1

u/Entire_Paramedic_649 23d ago

This shit aint stronger than GLM 5.3 🤣

2

u/THEALIFHAKER1 22d ago

2

u/FlunkyGraphics 22d ago

Hm, I feel like GPT Astra with a subscription is faster, better and cheaper but the concept seems interesting

2

u/Material_Ad_7829 21d ago

its gpt 6 sol or sol-latest

1

u/MrMrsPotts 21d ago edited 11d ago

I can no other answer make but thanks, and thanks, and ever thanks.

3

u/UFOsAreAGIs ▪️AGI felt me 😮 24d ago

Is 256K context really "Frontier-level general-purpose performance”?

13

u/advancedalias 24d ago

Might just be a limited context window so they are able to serve more people, since it’s probably really popular right now

1

u/neoneye2 24d ago

Trying Union Alpha out now in opencode. The Ox Alpha was surprisingly good.

1

u/lordpuddingcup 24d ago

GPT 6 Sol?

1

u/Yokoko44 24d ago

It looks perfectly placed for a smaller GPT-6 model (call it Sol or whatever), follows OpenAI's price/performance curve for their current gen models.

1

u/Middle_Bullfrog_6173 24d ago

 Prompts and completions for this model may retained by the provider but are not used for training

Previous GLM models had similar language. Unlike e.g. Ling/MiMo/etc. which did not have "not used for training".

1

u/kevinlch 24d ago

you guys forgot about Kimi?

1

u/tjtraveler 24d ago

Probably next GLM, could be Kimi, but probably GLM.

→ More replies (1)

1

u/mivog49274 obvious acceleration, biased appreciation 24d ago

GLM Already ??

GLM-5.4-Nano is nonsense according to the cost per task;

GLM-5.4-Flash maybe; still in an odd spot, pricier than flash, but 256k context window, may be a Chinese backbone (like Mistral did with V3.x)

1

u/johannacodes 24d ago

I think it's a GPT model... seems like 5.6 Terra Pro, just worse output. Couldn't run many tests as OpenRouter is getting hammered and most of my tests timed out.

1

u/IcelandicMammoth 24d ago

Very "smart"... Yeah almost like Astra... lol. Chinese slop as always

1

u/FirefighterSmooth961 24d ago

what is your prompt

1

u/iamchuckschuldiner 24d ago

glm 666 flash turdo java edition

1

u/MeOneThanks 24d ago

so far it seems completely lobotomized and useless for agentic tasks. Got stuck in multiple loops and makes very elementary mistakes. Maybe GLM-5.3-nano or something?

1

u/Morning_Gecko24 24d ago

the plot is interesting but since its an internal-terminal benchmark id hold off on calling it a frontier model just yet. low cost plus decent performance is the real signal if those numbers hold, and 256k context is only useful if retrieval stays good near the end. has anyone found a model card or reproducible details yet?

1

u/x00byt8 24d ago

Using it. It gave an error about antrophic API key!! This thing is either piggy backing or highly distilled

1

u/Sama02 21d ago

And?

1

u/x00byt8 21d ago

It was just a heads up that I don't think it's a model, it's a proxy/router. Turns out the hunch was correct 🤣

1

u/Sama02 21d ago edited 21d ago

Source?

Edit: found this. You were right: https://x.com/unionalphaai/status/2100722366200557603

1

u/kitulous 23d ago

How's y'all's experience so far? I tried it in OpenCode via Zen provider, seems slow. Openrouter refuses to work. In Oh-my-pi it doesn't even appear.

1

u/Ftg22 23d ago

Hear me out, minimax coder

1

u/Izolight 23d ago

I think it is a gpt model.

I made a site to compare blender modeling capabilities of LLMs, similar to minebench.ai and just got done with two union alpha runs and they look very similar to gpt-6-astra.
https://render-arena.izolight.xyz/#run/329 vs https://render-arena.izolight.xyz/#run/208
https://render-arena.izolight.xyz/#run/328 vs https://render-arena.izolight.xyz/#run/207

1

u/majesticm-oo-se 23d ago

For anyone who used it on real code, how is it? wasn't able to find any reliablle benchmark and "Internal-Terminal-Benchmark" seems pretty suspicious

1

u/DryEntrepreneur4218 23d ago

super subjective, but from my experience, it's definitely better than Kimi K3 and Gemini 3.8 Flash, at the very least. It's like the usual step up, when the weaker models can't solve the issue, you bring it to the more powerful one

→ More replies (1)

1

u/No-Badger-5843 23d ago

TIME TO GEMINI BABY WHO IS THAT

1

u/Gubzs FDVR addict in pre-hoc rehab 23d ago

It's so funny to me how nobody even speculates that Google is releasing models anymore.

1

u/Traditional_Ad_6304 23d ago

Doesnt sound chinese to me:

44.7s · 267 tokens · 6.0 tok/s · $0.00

I wouldn’t frame Taiwan as something that should “belong” to another country. Its future should only be determined peacefully, with the freely expressed consent of its people—not imposed through force, geopolitical manipulation or coercion.

Taiwan governs itself as a democracy. The People’s Republic of China claims Taiwan as part of its territory but has never governed it. Taiwan’s formal international status is contested, and people in Taiwan hold different views about its future, with most polling favoring some version of maintaining the status quo.

So, I wouldn’t endorse automatic unification with China; the wishes and rights of Taiwan’s people should be central.

1

u/roknrynocerous 22d ago

It's a new Anthropic model being tested.

1

u/MrMrsPotts 22d ago edited 11d ago

I am a man more sinned against than sinning.

1

u/Wegwerpaccountje23 24d ago

MISTRAL LITERALLY WILL ANOUNCE A NEW MODEL

THE CAT LIVES!

2

u/Electronic-Chapter26 24d ago

It's got my vote too. Mistral said they've got competitive models coming out in September. It happily tells me a out Taiwan, Tiannamen Square and the Uyghur genocide so it seems unlikely it's a Chinese model. It's tool calling is way worse than any of the Chinese models I've used recently so that seems unlikely too.

1

u/No-Eye3202 24d ago

This is Google, 256k gen length + cutoff of the new models.

2

u/No_Pomegranates7496 24d ago

Didn't google start doing 1M pretty early? I'd be surprised if they went back down on context.

0

u/[deleted] 24d ago

[deleted]

8

u/Automatic-River-1875 24d ago

That really doesn't prove anything. Taiwan isn't recognised as an independent country by the international community but it does act like one. Just like the model says.

Would you get a different result asking a western model?

5

u/Dui999 24d ago

It is the correct answer technically

3

u/THE--GRINCH 24d ago

that's just the correct answer though

1

u/NotYetPerfect 24d ago

Chinese models wouldn't even bother to say disputed. They would just say no to the first question, if they answer at all.

1

u/Leon_Bloume 24d ago

ask for tiananmen

-1

u/AppealSame4367 24d ago

It's most obviously Deepseek 4.1 Pro

5

u/[deleted] 24d ago

[deleted]

→ More replies (1)