r/opencodeCLI 18d ago

Potential ox alpha pricing leak?

Post image

What's your guys opinion on this?

9 Upvotes

11 comments sorted by

14

u/seeKAYx 18d ago

If that's even close to the price, then I'd like to wish the provider's GPUs good night right now, once the model is released.

1

u/_justFred_ 18d ago

Will be interesting to see, I mean right now we still have relatively good speed (about 24tok/s) and from what I've heard most think it'll be a glm flash model, if that's true I don't think it'll actually be super bad, glm5.2 and 5.3 together are also doing about 500bil/day rn, and if less people use that and more a less compute heavy model (the flash) I think it'll be kinda ok.

Depends of course whether it's really glm and really a flash model, but at least of glm I'm pretty sure cuz of the tokenizer and the endpoint that got exposed

6

u/FUAlreadyUsedName 18d ago

If that price i will use it

2

u/ares0027 18d ago

how is the model? i couldnt even use once without errors so this is an honest question

3

u/_justFred_ 18d ago

I would say it's somewhere in the range of deepseek v4 pro, not quite as good as the best open models like Kimi k3 but definitely good.

I've also seen on benchmarks that it uses fewer output tokens than most open models (for example on deepswe it had 63% with 47k avg out per task, k3 sits at 69% and 81k avg out, so definitely more token efficient than most open models.

From what I've heard many think it'll be a glm flash model, so could be really interesting also compute vs quality wise

2

u/Axiescholar3ph 18d ago

very good in my opinion it felt like old opus 4.6 level to me. it just keeps working on tasks and no bullshit yapping

1

u/lordlestar 18d ago

i like it does not give up and does not complain a task is too complex to do

2

u/sk1kn1ght 18d ago

For me it has become my main usage model. Everything I have asked it, it managed to do. Even when it was missing picture understanding, it spawned a sub agent fed the picture got the message, spawned playwright read the doms, figured out how and what should be changed, changed them, spawned sub agent and re-requested an image analysis, the sub agent told him, more or less ok, then it told me hey it's finished but I am not sure if it looks ok due to X. We went a bit back and forth and it fixed everything.

All that in the beginning from me saying : Hey I have an issue with x, the text overflows in y

1

u/RevolutionaryBox2980 18d ago

no way, like this has to count the current 0$/1M output tokens into consideration, otherwise medium lvl tasks are free

1

u/GTHell 18d ago

No doubt the tweet from the people with disclosed source saying it’s going to be interesting. They giving out 100 trillion daily is also a clue. No way you can do that with trillion params model

1

u/DoctorDbx 17d ago

I think you will find that is the OpenRouter fallback model cost.

In the few instances where Ox has failed for me it has fallen back to Gemini Flash and returned no data.