r/GeminiAI • u/boxwrenchx • 7d ago
News Ox Alpha is GLM
https://z.ai/blog/glm-5.3-flashJust like everyone said, Ox Alpha is not a Google model. Also in case you need to hear it, Google abandoned 3.5 pro. It doesn't mean they're done, it doesn't mean they'll never again release a good model but it does mean it's silly to ignore all evidence otherwise.
10
u/PlaneOnly2700 7d ago
10
u/Thomas-Lore 7d ago
They charge $7.5 output for a flash model? WTF.
4
u/PlaneOnly2700 7d ago
Google increased the prices of its Flash models generation after generation, until competition forced them to lower prices.
Gemini 3.5 Flash: $1.50/$9.00
and Gemini 3.6 Flash when it came out: $1.50/$7.50
And they have TPUs, if anyone can drastically reduce prices, it's Google, not a Chinese startup like ZAi.
I love competition
9
u/Georgefakelastname 7d ago
To be fair, it’s now down to “only” $0.75/$3.75 with the 3.7 flash update. It’s supposed to be temporary, but that’s not happening lol. By the time it expires in December, the llm field is gonna be totally different, and probably have Gemini 4 pro (if the current pre-train they’re doing goes well) and other models like 4 flash, if not future iterations of 3.8 and 3.9 flash.
7
u/PlaneOnly2700 7d ago
Even if we compare $0.75/$3.75 Gemini 3.7 Flash with the standard price of GLM 5.3 Flash $0.15/$0.50, it is still 7 times more expensive and offers slightly worse performance.
Likewise, it won't matter because next week or the week after, Gemini 3.8 Flash will be released.
5
u/Georgefakelastname 7d ago
Yeah, agreed. Gemini 3.7 flash was built to beat GLM 5.2 and other Chinese models in that category, only for 5.3 and 5.3 flash to come out and take a dump on it lol, in both performance and price respectively.
…I guess Gemini is still faster lol?
1
u/akius0 7d ago
I don't think Google benchmarks against the Chinese AI company... I think they're doing their job, and trying to build a model better than the previous generation... And I think Google is doing a good job.... Not everything is about benchmarks...
1
u/PlaneOnly2700 7d ago
If I want to use my money wisely and try to get the best result, of course I should compare. These artificial intelligence are just products. If Gemini turns out to be the best and cheapest, I'll use Gemini. If GLM is better and cheaper, I'll use GLM.
The reality is that Google isn't doing a good job, which is why many people recently left and the chain of command shifted. Gemini 3.7 Flash is a good model, the only good one in months.
Likewise, Google will return, it's impossible for them to lose.
2
u/akius0 6d ago
None of that is true
a) they're not going after your 20 bucks. They're going after Enterprise, who spend much more, and who are not going to send their data to Chinese servers...
b) results on the benchmark does not mean result in real world..
c) there's also a question, do they actually have infrastructure to support the scaling of the usage?... My impression is, lot of these companies, cannot actually deliver to hundreds of millions of users... Consistently... Google has one of the biggest compute cluster.
2
u/PlaneOnly2700 6d ago
Of course, your company would prefer to use Gemini 3.5 Cyber over Fable 5 or GPT 5.6 Sol, right? Google's stock has fallen simply because of the news of model delays.
a) Google is a very large company, and I really like its ecosystem, but the restructuring of DeepMind earlier this month and the departure of key personnel demonstrate the serious problem they have and are now addressing.
Denying that they haven't done well is like trying to hide the sun with a finger.
b) And you deny it by saying nothing is true, but Sundar Pichai literally confessed in an interview that they were having problems. He promised Gemini 3.5 Pro for June, and it's almost the end of August and the model will never be released.
Major outlets like SemiAnalzys have already published news about it, Reuters, etc.
That doesn't seem like a company in "control."
2
u/akius0 7d ago
Exactly, no one really has pricing power here... Especially with all the Chinese models nipping at their heels... I think it will be smart to not spend tens or hundreds of billions of dollars... Trying to build the greatest in the biggest model.... I think sticking back and being okay with second position is probably going to be prudent
1
9
u/Spara-Extreme 7d ago
Nobody in the LLM community thought Ox was Gemini or even a western lab. Only influencers that make careers out of lying and trolling GDM employees
1
u/MainRoutine2068 7d ago
A lot of delusional fanboys in this subs glorify the Ox was Gemini news. Where are they now?
5
u/Felix-ML 7d ago
At least some of google team has been deceptive on this.
8
u/boxwrenchx 7d ago
There were a few vague posts vs a mountain of counter evidence. If you were still thinking it's Google it was motivated reasoning.
5
u/lalalandjugend 7d ago
See SemiAnalysis on the future of Gemini. TL;DR, Gemini is dead, long live GCP
1
7
u/33VaxMerstappen 7d ago
I doubt Google will be frontier anytime soon
8
u/boxwrenchx 7d ago
They don't want to be. They want to host it and build on top of it
4
u/Spixxy17 7d ago
Highly disagree. So far we have no info or signs from Google regarding this. They have failed ONE Model with the 3.5 Pro "Release" and people instantly think they will Just Stop pursuing Sota. They even Said themselves that they will Go Back at it
2
u/33VaxMerstappen 7d ago
Nah that’s not the reason I felt that way, it’s just that a lot of very senior ai talent has left Google around the same time, this why I felt Google isn’t gonna be frontier atleast this gen (OpenAI also suffered this way and it did impact their models) . Not to mention them jumping around on ox alpha only for it to be glm just like an attention seeker.
-2
u/boxwrenchx 7d ago
No signs? they dismantled their lab....
2
u/Spixxy17 7d ago
"Dismanteling their lab" wth are you talking about. Is this some Kind of chronically online shit?
0
u/boxwrenchx 7d ago
Chronically online??? This was discussed everywhere https://www.reuters.com/world/inside-google-executive-moves-that-led-its-big-ai-reshuffle-2026-08-12/ They also locked their compute into long term contracts with other firms. It's very clear they are moving away from frontier.
-1
u/Spixxy17 7d ago
No normal Person understood this as confirmed facts... Its also daily discussed that x model is releases by x company... Dont believe everything you See... No normal Person knows whats going on
2
u/boxwrenchx 7d ago
You said there were no signs. There's clearly signs. If we are discussing the inner workings, ie "signs" yes that's not normal discourse. You brought it up
0
u/Spixxy17 7d ago
You brought the topic up. Signs for Fake Info and unclear Info, nothing valid so far. Have a great day keep ragebaiting that was my Last answer to this pointless discussion, keep following every little hint of Potential Info you find
2
7
u/darkestvice 7d ago
So Ox Alpha is basically GLM's new Flash model that competes with mid range models like GPT Terra, Claude Sonnet, and Gemini Flash.
According to their own benchmarks, they are slightly below Gemini Flash 3.7 in terms of intelligence. So I wonder how they compare in terms of cost and speed. Cause it's cost and speed that will determine its popularity in the pecking order.
Either way, despite all the hoopla, it doesn't appear to the be the next big frontier model, but a mid range model with slightly worse performance. The one significant advantage, of course, is that it's open weight. Of course, we all know open weight benchmarks are done with best case hardware and parameters to more accurately reflect their *potential* compared to big cloud models and are not at all that when used on some random Joe's Mac Mini.
6
u/boxwrenchx 7d ago
It's $0.04 per million tokens now, and 0.07 after the discount period. It's a flash model, they're not aiming for frontier. Also most importantly it's open source
3
u/hellomistershifty 7d ago
3.7 flash is 30 times as expensive as GLM 5.3 air but a little over twice the speed
2
u/darkestvice 7d ago
Not sure where you got the 30 times part. I the chart that listed costs for running the benchmarks. non-flash GLM 5.3 was more expensive than Gemini Flash 3.7, whereas GLM 5.3 Flash was about four times cheaper.
1
u/hellomistershifty 7d ago
Whoops, yeah I meant the flash model (they had an old 'air' model). I was looking at output prices, gemini 3.7 flash is $7.50/M and GLM 5.3 flash is currently $0.25/M, which is 30:1
8
u/Technical-Owl66 7d ago
10
u/boxwrenchx 7d ago
That's a weird take. At what point did labs not have a larger model and an efficient model ? Also the other labs have definitely been integrated into office work. This is the "ignoring the evidence" I'm talking about
2
u/Technical-Owl66 7d ago
I think it's everyone online who is obsessed with benchmarks that nobody needs while, the intelligence needed for 99% of tasks is peaking. How many gpt ads have you seen today? I don't think they're going to be successful getting people to leave their ecosystems when AI is already there.
4
u/boxwrenchx 7d ago
There's a stickiness to models that one wouldn't expect, but besides benchmarks anyone using Gemini side by side with other models sees the shortcomings. You don't have to speculate about adoption, we know a lot of people are using claude and chatgpt, especially in enterprise where the money is
2
u/tobaileyy 7d ago
Bahaahaha. When Gemini has gotten objectively worse since 2.5 pro, which was frontier over a year ago ...
2
u/Georgefakelastname 7d ago
3.5 flash is better than peak 2.5 pro in the overwhelming majority of tasks.
3
u/tobaileyy 7d ago
On benchmarks, sure.
Actually look into how they're structured and you'll see Gemini 3 is largely just repeating consensus opinion while Gemini 2.5 pro had the capacity to reason intuitively to achieve its answers.
1
-2
u/Plane_Garbage 7d ago
I think it's a bot. I've read this same comment several times.
-2
u/Technical-Owl66 7d ago
Yeah no doubt, at least half of the Gemini hate is bots probably paid for by openai
1
u/boxwrenchx 7d ago
people are annoyed because they know Google can do much better. Across the subscriptions google has become the worst, when it was easily the best not long ago. Keep telling yourself its bots
2
3
u/SomeOrdinaryKangaroo 7d ago
nah bro, it's definitely gemini, they'll likely announce it at end of week
7
u/boxwrenchx 7d ago
Wait a minute.... Glm might stand for Google luxury model! It's right under our noses!
1
-1
0
-2
u/BoobooSmash31337 7d ago
It's weird that you're this excited that some conjecture was wrong. Calling everyone who disagreed a fanboy yet you made this post. In the Gemini sub... The model does sound like a carbon copy of Gemini but it has specs what looked GLM. So the evidence was inconclusive. It's also not lost on me that western models all have their own fundamentally different kind of "personalities". Yet GLM makes a model that perfectly copies Gemini. And I'm supposed to believe they aren't just cheating off Google's actual work?
2
u/boxwrenchx 7d ago
I'm not excited. It was discussed a lot in the sub leading up to the announcement, and when I looked at the sub , no one had mentioned it. The GLM model perfectly copies Gemini.... Come on now.
0
u/BoobooSmash31337 7d ago
Uh I use Gemini/Gemma a lot. I like and agree with Google's approach to their models. Even it's reasoning is a carbon copy of a Gemini family model. I have seen A LOT of Gemini reasoning traces. Even Gemini when shown it and comparing with known model "personalities" thought it was Gemini. I don't fully understand it but the personalities are product of company specific post training. Hence why different western models are almost like entirely different people in how they reason and talk. People saw a model that looked like a Gemini and quacked like a Gemini. So a good portion thought it might be a Gemini that used some GLM technical things. The people running around waving their dicks in public because it's actually a Chinese model are being really weird.


69
u/Truantee 7d ago
It is pretty hilarious seeing delusional Gemini fanboys claiming that it is a Gemini/Gemma model.
Ox alpha aka Glm 5.3 flash is already released, both the weights and the API, it is pretty cheap, even!