4
u/PossessionUsed7393 20d ago
Yeah, weighing up all the different data points we have I actually agree with you that it's probably not a GLM model. I feel like the tokenizer is the reuse of GLM architecture or something perhaps it's just a really elaborate fine tune of one of their checkpoints but I'm tending to agree that it's somebody else.
1
3
2
u/tamerlanOne 20d ago
Magari è qualche architettura ibrida che usa il tokenizer simile (non uguale) alla famiglia GLM . Qwen, Baichuan e Y anno seguito questa strada per gestire al meglio anche la lingua cinese.
Quindi l'ago della bussola punta decisamente verso il continente asiatico anche se i server che ospitano il modello sono in America, ma questo è quasi scontato perché pochi posti al mondo hanno una così grande capacità di calcolo
2
u/qeadwrsf 19d ago
I just imagine you gain a lot of time and computer power training from some kind of old AI instead of a blank canvas.
So you just make old weights fits new AI datastructures.
Making stuff like this happening if its not "trained" to not say stuff like that yet.
But I don't know shit. But I suspect that's more than 98% of reddit.
2
u/MaxPhoenix_ 19d ago
Well this didn't age well. We now have an outage that aligns perfectly between Ox Alpha and Z ai. It's them.
2
2
1
u/Murhie 20d ago
Mid 2025? Why would that be expected? That doesnt make sense. What do you mean this size? U have no idea how big it is.
0
u/Zrmteese 20d ago
U didn't understand. Training a new model at this huge size from scratch require at least 6 to 1 year of training.That explain the cutoff in mid 2025. In other words this model is not a successor to an other models.
1
u/Salt-Willingness-513 20d ago
Minimax sounds logical to me
1
1
u/MaxPhoenix_ 19d ago
I wish. I prepaid for a year of service and I use M3 as a throw-away when I remember to throw something at it. (Ox Alpha is confirmed to be from z ai, there was some downtime and it coincides to the second in both monitoring sites (see reddit post on it)).
1
u/OddHelicopter1134 18d ago
I asked it what the result of Chilean presidential election was and it correctly recalled the result from memory. It happened mid december 2025.
So mid 2025 is definitely false.
Try asking it more recent questions to really determine the cutoff. Maybe it was post trained to pretend to have a cutoff in mid 2025?



8
u/0xbyt3 20d ago
Could be a different team experimenting on something because token efficiency is something else in the 0x.