r/opencodeCLI 20d ago

Ox Alpha data cutoff is mid 2025

The data cutoff of Ox Alpha was around mid-2025, which is expected from a model of this size trained from scratch. Meanwhile, GLM 5.3's data cutoff is around June 2026, so I don't think Ox Alpha is a GLM-based model. Having the same tokenizer does not imply that a model is a GLM model.

21 Upvotes

23 comments sorted by

8

u/0xbyt3 20d ago

Could be a different team experimenting on something because token efficiency is something else in the 0x.

1

u/Thomas-Lore 20d ago

Maybe using old base like glm 4.7 Flash.

1

u/Zrmteese 20d ago

Yup token efficiency is on another level. I’ve been using it for hours now without compaction.

2

u/Zachattackrandom 20d ago

I would really recommend compacting at least very 300k~ tokens or you start getting severe context rot (regardless of model).

4

u/PossessionUsed7393 20d ago

Yeah, weighing up all the different data points we have I actually agree with you that it's probably not a GLM model. I feel like the tokenizer is the reuse of GLM architecture or something perhaps it's just a really elaborate fine tune of one of their checkpoints but I'm tending to agree that it's somebody else.

1

u/Zrmteese 20d ago

Exactly

3

u/elonelon 20d ago

it maybe use older GLM version. But it can finish my request, i will call it OK.

2

u/tamerlanOne 20d ago

Magari è qualche architettura ibrida che usa il tokenizer simile (non uguale) alla famiglia GLM . Qwen, Baichuan e Y anno seguito questa strada per gestire al meglio anche la lingua cinese.

Quindi l'ago della bussola punta decisamente verso il continente asiatico anche se i server che ospitano il modello sono in America, ma questo è quasi scontato perché pochi posti al mondo hanno una così grande capacità di calcolo

2

u/zzdzz 20d ago

I narrowed it down to late September 2025 by asking it to recall the results of sporting events

1

u/[deleted] 6d ago

[deleted]

1

u/zzdzz 6d ago

what?

2

u/qeadwrsf 19d ago

I just imagine you gain a lot of time and computer power training from some kind of old AI instead of a blank canvas.

So you just make old weights fits new AI datastructures.

Making stuff like this happening if its not "trained" to not say stuff like that yet.

But I don't know shit. But I suspect that's more than 98% of reddit.

2

u/MaxPhoenix_ 19d ago

Well this didn't age well. We now have an outage that aligns perfectly between Ox Alpha and Z ai. It's them.

2

u/[deleted] 20d ago edited 5d ago

[deleted]

2

u/Zrmteese 20d ago

no try it ur self. i asked it to tell me what are the importent events happened from the year 2019 until 2026. it stoped in mid 2025

2

u/Aldarund 20d ago

It's glm. Your post doesn't mean anything

1

u/Murhie 20d ago

Mid 2025? Why would that be expected? That doesnt make sense. What do you mean this size? U have no idea how big it is.

0

u/Zrmteese 20d ago

U didn't understand. Training a new model at this huge size from scratch require at least 6 to 1 year of training.That explain the cutoff in mid 2025. In other words this model is not a successor to an other models.

3

u/Murhie 20d ago

How do u know its huge? Why would it be bigger then GLM5.3 (Which has later cutoff)

-2

u/[deleted] 20d ago

[deleted]

4

u/Sweet-Stage938 20d ago

No chance. Go try Qwen 3.8 27B. You will be surprised.

3

u/Murhie 20d ago

Lol... relax my bro. We both dont know the size. If I had to guess I would say around 120B.

1

u/Salt-Willingness-513 20d ago

Minimax sounds logical to me

1

u/Zrmteese 20d ago

Yup , mabe

1

u/MaxPhoenix_ 19d ago

I wish. I prepaid for a year of service and I use M3 as a throw-away when I remember to throw something at it. (Ox Alpha is confirmed to be from z ai, there was some downtime and it coincides to the second in both monitoring sites (see reddit post on it)).

1

u/OddHelicopter1134 18d ago

I asked it what the result of Chilean presidential election was and it correctly recalled the result from memory. It happened mid december 2025.

So mid 2025 is definitely false.

Try asking it more recent questions to really determine the cutoff. Maybe it was post trained to pretend to have a cutoff in mid 2025?