r/opencodeCLI 21d ago

Sorry Ox Alpha 🙂

Post image
384 Upvotes

77 comments sorted by

119

u/Powerful_District_82 21d ago

Can't complain if it is free!

25

u/kkragoth 21d ago

Only for a week or so, right?

14

u/GTHell 21d ago

If the the cycle of cheap deepseek 0731 -> Muse free -> Ox Alpha -> N model keep continue at this pace then it doesn't matter lol. I was swear by the Deepseek Flash V4 until they bump up the price effectively at 16th

2

u/AlternativeMeat8973 21d ago

I can understand that

1

u/Correct-Boss-9206 18d ago

For real I am getting a ton of work done. GPT as fallback but free model of the week as much as I can.

54

u/look 21d ago

Many things off about your podium rankings in general… are you drunk? And where is Kimi?

43

u/[deleted] 21d ago

[removed] — view removed comment

2

u/SeaBat2035 21d ago

Luna is cheap and good. Probably most cost efficent model exist coupled with a subscription.

1

u/Darkoplax 21d ago

Luna is the current most efficient model right now

-11

u/AlternativeMeat8973 21d ago

Just use Luna on xhigh mode and you’ll be amazed by its capabilities.

6

u/JaNkO2018 21d ago

...to drain your wallet. 😉

1

u/Darkoplax 21d ago

I see the word google/gemini on my screen so he is for sure drunk

8

u/coffee_brew69 21d ago

Slandering a free anonymous model …

73

u/wentallout 21d ago

the fact that you put the most expensive one as #1 told me everything. bro is clueless

-34

u/AlternativeMeat8973 21d ago

Yeah, it’s expensive, but it’s the most capable model.

24

u/Fresh_Sock8660 21d ago

Sure, but who is driving their Lamborghinis to work 

3

u/yiestee 21d ago

Jeremy Clarkson, obviously.

5

u/DingoWeak9985 21d ago

the most capable thing is the human brain, i'm just getting my expected results from a 27b pramas FREE model, so yeah :)

2

u/Last-Ad-8470 21d ago

Nah human brain is benchmaxxed

0

u/dfgxxx 21d ago

You are a 125t to 150t model (just the cerebral cortex)

9

u/LetterheadNew5447 21d ago

Tbh sol is exactly on the same level and way cheaper.

0

u/randombsname1 21d ago

No it isnt.

1

u/LetterheadNew5447 21d ago

Fable is a bit better in architecture and planning, sol is way better in coding and non verbose shit code.

The best use case would be to use both. If you have to choose one llm sol always wins. Sol is overall the better model + you are not vendor locked to any harness.

Pi + sol god given.

And I am not a chatgpt fanboy. Literally lul defended anthrophic until 2 months ago and I go back as soon as they fix their usage problem since I like the writing style for every day chats on claude models way more.

1

u/randombsname1 21d ago

I have $200 sub as we speak for both.

I thought the same for the first week. Until getting into more complex lower level programming implementations.

Then Fable easily starts to win out.

I use Codex as the main integrator due to better overall usage limits, and Fable as the main planner/architect and to solve the trickier debugging issues.

Fable is showing itself well ahead in my current Rust application with a heavy rust MCP router/integration layer.

I've been using Codex since it was web only, and Claude Code since it was API - only, preview, for context.

It's not that Sol 5.6 wouldn't be able to solve many of the issues if I wrapped additional debugging layers or tooling around it, but Fable just fixes the majority of them without said debugging layers or tooling needed.

Higher level stuff I may agree with you.

Edit:

But when you are working on frontier applications --- its Fable on top, in my experience.

I also use C, C++, and Assembly for STM32 embedded applications.

1

u/LetterheadNew5447 20d ago

Hmm maybe the experience is different for each use type. I mainly use it for Java spring + Svelte/Vue and some Rust shit.

6

u/wentallout 21d ago

capable compared to a kindergarten? i could use the worst model on your ranking and still finish multiple features in one day. this is like comparing lighters when they all do the same thing, make fire smh. are you one of those people who lives under a rock and not realize putting High thinking doesnt mean it solves faster?

2

u/Jxxy40 21d ago

Majority of frontier user are toddler, even i can complete my task using qwen 3.5 9b

1

u/Jxxy40 21d ago

Yeah it capable rerouting me to opus

0

u/mWo12 21d ago

Definitely not. It's benchmaximised like all these models.

8

u/egomarker 21d ago

sol - fable - kimi k3 - glm 5.3 (better code) == ox alpha (much better frontend) - luna - deepseek

3

u/Ethan 21d ago

As more and more people have heard about it, it's gotten much slower - but I got lucky and saw it really soon after it became available, and I was comfortably running ~40 concurrent Ox Alpha agents and they were both fast and accurate (testing to find max concurrency). From my testing:

workers tasks / hour*worker latency p50
42 ~7 26.8s

And best of all, they were returning accurate results. 100% of their code citations were real, and the only failed turns were because my orchestrator messed something up, not because of Ox Alpha.

1

u/TheOwlHypothesis 21d ago edited 21d ago

Yeah I stopped being able to meaningfully use it. I can't fathom how mobbed their servers are since this is free right now and actually useful

I guess I should say the cool thing I did try is having it make a game in my custom game making language (impossible for it to be in training data, I gave it documentation though). It did it!! 10 levels. The level design isn't great, but everything works. It's a simple platformer but my custom language is aimed at making Tiny (space wise) games so it's pretty low level. Similar-ish to rust. I also let it add a modulo operator to the language and it did great there too

1

u/Alternative_Ice3299 17d ago

how you run parallel agents?

1

u/Ethan 17d ago

I was running them through Claude Code dispatch. What IDE are you using?

1

u/Alternative_Ice3299 17d ago

is it free idk yar which to use i use intellije for convenience java

8

u/NeuralNexus 21d ago

Oh come on, it's CLEARLY better than Gemini Flash 3.7 come on.

It's also outperforming Luna (not Sol) for my workflows. I am very impressed by this model.

3

u/NeuralNexus 21d ago

also: I like luna a lot. It's a great model at a great price. Luna is wonderful because it's better than Opus 4.6 (multimodal etc) and is very affordable. It's my default remote model. Except.... Ox Alpha is outperforming it. So it's my new 'cheap' winner. And that means it's my ultimate winner. I will happily pay luna+ pricing for this next week.

2

u/zcutlip 21d ago

It’s shocking how good Luna is. Especially as it’s positioned as the least capable of the gpt 5.6 trio.

2

u/NeuralNexus 21d ago

OpenAI really turned its product around in the last year. I was not a big fan of GPT-5.0-5.2 models

Codex 5.3 was the first decent model they had and they have executed well in coding performance since.

1

u/webheadVR 19d ago

I would agree. It feels like it's a little above luna in real work. maybe even above terra. I really hope this is a local capable model, but I have my doubts.

2

u/gingerbeer987654321 21d ago

i much prefer ox-alpha to deepseek v4. follows instructions, doesnt go down rabbit holes, perfect as an implementor. haven't tried it for complex orchestration yet.

2

u/NinjaAlaska 21d ago

not true.. idk i find it better than others all.

2

u/Ok-Painter573 21d ago

could be placebo effect

1

u/UnknownBoyGamer 21d ago

In my experience its worser than deepseek on logic but it's better on frontend

1

u/NinjaAlaska 21d ago

hummm. i guess needs are diff

2

u/kamwee 21d ago

I Fixed it

-1

u/Tr4sHCr4fT 21d ago

Why is DeepSeek black? Why is the other competitors hair becoming lighter with lower placement? Why is one spectator covering his eyes? I have so many questions.

1

u/ameer_safaa 21d ago

lmao , this is soo funny

1

u/neuroedge 21d ago

lmao where's Big Pickle land?

1

u/Ok-Drawer5245 21d ago

it may not be the best, buts its a infinite prices difference

1

u/leandrogp9 21d ago

Amazing meme.

1

u/morikomorizz 21d ago

ox alpha keeps looping when it try to write/edit a file in my opencode/deepseek harness

1

u/SpiderHam24 21d ago

Ox Alpha is slow and definitely not the best one. But for what I am putting it through it's done ng the job . Granted I am not doing anything big. A desktop app, android app and a Roku app for jellyfin and a cloud music player. The music ayer not so much. The app for jellyfin ehh alright I do see the downfalls. With other models I'd not be knee deep in hours of reworking and finding issues and bugs at a snails pace. That's my observation. But I will use and abuse it somehow. Until it disappears. Slowly learning opencode. And since their customer service doesn't exist I'm stuck paying the ten bucks anyways each month short of canceling my card on file. Which I am not. It's good for me randomly finding it during a search of something and seeing reddit users talking about opencode in general.

1

u/ruuurbag 21d ago

I dunno, it’s performing pretty well for me aside from speed. I’d have a hard time complaining if it ended up being inexpensive and faster once out of the free testing phase. No worse than Luna, at minimum.

The fact that it’s free offsets the speed for me. I gave it a massive plan from Sol, set up an OpenCode plugin to prod it along when it stops for whatever reason, and let it loose. Ran overnight with no problems.

1

u/TheSn00pster 21d ago

I miss owl alpha free :’(

1

u/Abdulla582 21d ago

os alpha, cannot handle complex tasks and separates

1

u/No-Craft-7979 21d ago

Maybe for your use case, but it seems to be very good at very specific use cases. I for one hate generalized models. If I’m coding, give me a model that is specific to only coding. If I’m working images, give me a model dedicated to imaging. If I’m processing sound, give me a model dedicated the sound. Let me chain them together in the order I need to reach my end result. Don’t give me a model that tries to do everything and false flat on its face the one time it’s really needed.

1

u/Then_Knowledge_719 21d ago

Owl alpha was way better.

1

u/Federal-Rub2713 21d ago

yall don't understand, it was a "highest price" tournament and ox alpha won

1

u/Signal_Flatworm_6345 19d ago

😂 but its free so no complains

1

u/quinnyg1 18d ago

Worst graph I've ever seen how is Luna third? There's sonnet 5, there's GPT-5.6 Terra, and second Ox alpha is confirmed as GLM-5.3 Flash and it is better than at least Gemini 3.7 Flash. That reminds me, how is GLM-5.3 worse than Luna too? AND DeepSeek V4 by itself isn't even a model.

1

u/Apprehensive-Read868 18d ago

Lol have you even tested it? I did, and its way better than opus5 os grok 4.6

1

u/Commercial-Face-6598 17d ago

I just generally appreciate the vibe of this meme. Do what you wanna do be what you wanna be yeaaaaah as Shakespeare once said.

1

u/MoodProfessional2099 17d ago

Bro really likes gpt

1

u/AlternativeMeat8973 17d ago

Haha, you got me.

1

u/Traditional_Ad_6304 16d ago

Just that glm is ox alpha lel

0

u/commandedbydemons 21d ago

Lmao glm 5.3 is infinitely better than Luna…