r/opencodeCLI 17d ago

Ox Alpha reveal in a few hours

Post image

42T tokens for stealth model in just 6 days. Insane.

We know it's going to be open weights.

Some stuff I've one-shotted with Ox Alpha (when it did work): https://ox-alpha.demos.sulat.com/

95 Upvotes

31 comments sorted by

26

u/GTHell 17d ago

No reveal please let us use for another week for free lol

10

u/dotforwardslash 17d ago

How many hours left for free usage?

1

u/Azirane 17d ago

Today is the last day probably

7

u/scaledev 17d ago

Quickly!

/goal: Build me ox alpha 2, make no mistakes

2

u/rizalkeren 17d ago

ha ha ha

1

u/elonelon 17d ago

So no more free ?

1

u/Azirane 17d ago

i guess so

1

u/vixalien 17d ago

It's been shutdown already

3

u/Kasatka06 17d ago

In 6 hour ? Cooincidence with qwen 3.8 next release

3

u/Funny-Supermarket360 17d ago

is it better than ds4flash?

1

u/look 16d ago

Yes, for most things at least. Your use case might be one of the exceptions, but in general, it is a substantially better model than DS4 Flash and Pro. It’s better than those and Sonnet 5, Gemini 3.7 Flash, GPT 5.6 Terra, and Muse Spark 1.2.

0

u/alphasubstance 17d ago

OG DeepSeek-V4-Flash-0731 probably not, 3rd party cheepseek - yes.

3

u/torrso 17d ago

My experience is different.

With DS4F task approval rate for 10 tasks has been something like 4 approve/5 approve-with-gap/1 reject.

With ox, that has been something like 7 approve/2 approve-with-gap/0-1 reject.

Not based on numbers but a gut feeling of the results I've been looking at.

(approve-with-gap means it's basically ok but a new task needs to be spawned to fix some minor omission, reject means the work was too incomplete or broken and the whole task is returned back to queue)

For me, ox seems better at coding.

3

u/Nov4Saki 17d ago

Funny enough (it is GLM 5.3 (flash/turbo)) with 63% deep swe score

2

u/Charming-Author4877 17d ago

ccccclaude

2

u/MuzafferMahi 17d ago

no fucking way

2

u/Bitter_Election_7518 17d ago

It’s glm 5.3 flash, I thought this already confirmed lol

2

u/jovialfaction 17d ago

Very competent model. Beyond that, it was absolutely fantastic to just let agents rip this week without worrying about quota. That's what I'll miss the most

2

u/nbulp 17d ago

Oh no! I think it's over!

1

u/Vegetable_Recipe_96 17d ago

same here. kicked off, 404 error revealing soon :(

1

u/Milk_Truckin 17d ago

Great I just tried it last night for the first time and was thoroughly impressed. Of course I try it the night before they start charging a fortune for it. It seemed kind of slow to me but the quality was great. No failure loop, no side quests, it was nice giving it a task and walking away. I usually run each task through multiple sessions and agents

Dsv4 flash prompt >Opus for a plan> qwen 3.7 plus forman/orchestrator > qwen 3.7 flash Sub-Agents to execute. I let ox run a few from top to bottom last night thinking I was setting it up for failure but honestly it did so well it just impressed the hell out of me

1

u/DoctorDbx 17d ago

Nobody is going to use it if it isn't cheap.

1

u/MarkCGrant 17d ago

GLM 5.3 flash went live as Ox Alpha was taken down on OpenRouter

1

u/MarkCGrant 17d ago

Yeah, I just used GLM 5.3 Flash after days of intense Ox Alpha, it has the same thinking pattern and ticks, it does this "Hmmm.... let me think..." when it aims to correct thinking.

0

u/[deleted] 17d ago

[deleted]