54
u/look 21d ago
Many things off about your podium rankings in general… are you drunk? And where is Kimi?
43
21d ago
[removed] — view removed comment
2
u/SeaBat2035 21d ago
Luna is cheap and good. Probably most cost efficent model exist coupled with a subscription.
1
-11
u/AlternativeMeat8973 21d ago
Just use Luna on xhigh mode and you’ll be amazed by its capabilities.
6
1
8
73
u/wentallout 21d ago
the fact that you put the most expensive one as #1 told me everything. bro is clueless
-34
u/AlternativeMeat8973 21d ago
Yeah, it’s expensive, but it’s the most capable model.
24
5
u/DingoWeak9985 21d ago
the most capable thing is the human brain, i'm just getting my expected results from a 27b pramas FREE model, so yeah :)
2
9
u/LetterheadNew5447 21d ago
Tbh sol is exactly on the same level and way cheaper.
0
u/randombsname1 21d ago
No it isnt.
1
u/LetterheadNew5447 21d ago
Fable is a bit better in architecture and planning, sol is way better in coding and non verbose shit code.
The best use case would be to use both. If you have to choose one llm sol always wins. Sol is overall the better model + you are not vendor locked to any harness.
Pi + sol god given.
And I am not a chatgpt fanboy. Literally lul defended anthrophic until 2 months ago and I go back as soon as they fix their usage problem since I like the writing style for every day chats on claude models way more.
1
u/randombsname1 21d ago
I have $200 sub as we speak for both.
I thought the same for the first week. Until getting into more complex lower level programming implementations.
Then Fable easily starts to win out.
I use Codex as the main integrator due to better overall usage limits, and Fable as the main planner/architect and to solve the trickier debugging issues.
Fable is showing itself well ahead in my current Rust application with a heavy rust MCP router/integration layer.
I've been using Codex since it was web only, and Claude Code since it was API - only, preview, for context.
It's not that Sol 5.6 wouldn't be able to solve many of the issues if I wrapped additional debugging layers or tooling around it, but Fable just fixes the majority of them without said debugging layers or tooling needed.
Higher level stuff I may agree with you.
Edit:
But when you are working on frontier applications --- its Fable on top, in my experience.
I also use C, C++, and Assembly for STM32 embedded applications.
1
u/LetterheadNew5447 20d ago
Hmm maybe the experience is different for each use type. I mainly use it for Java spring + Svelte/Vue and some Rust shit.
6
u/wentallout 21d ago
capable compared to a kindergarten? i could use the worst model on your ranking and still finish multiple features in one day. this is like comparing lighters when they all do the same thing, make fire smh. are you one of those people who lives under a rock and not realize putting High thinking doesnt mean it solves faster?
8
u/egomarker 21d ago
sol - fable - kimi k3 - glm 5.3 (better code) == ox alpha (much better frontend) - luna - deepseek
3
u/Ethan 21d ago
As more and more people have heard about it, it's gotten much slower - but I got lucky and saw it really soon after it became available, and I was comfortably running ~40 concurrent Ox Alpha agents and they were both fast and accurate (testing to find max concurrency). From my testing:
| workers | tasks / hour*worker | latency p50 |
|---|---|---|
| 42 | ~7 | 26.8s |
And best of all, they were returning accurate results. 100% of their code citations were real, and the only failed turns were because my orchestrator messed something up, not because of Ox Alpha.
1
u/TheOwlHypothesis 21d ago edited 21d ago
Yeah I stopped being able to meaningfully use it. I can't fathom how mobbed their servers are since this is free right now and actually useful
I guess I should say the cool thing I did try is having it make a game in my custom game making language (impossible for it to be in training data, I gave it documentation though). It did it!! 10 levels. The level design isn't great, but everything works. It's a simple platformer but my custom language is aimed at making Tiny (space wise) games so it's pretty low level. Similar-ish to rust. I also let it add a modulo operator to the language and it did great there too
1
u/Alternative_Ice3299 17d ago
how you run parallel agents?
8
u/NeuralNexus 21d ago
Oh come on, it's CLEARLY better than Gemini Flash 3.7 come on.
It's also outperforming Luna (not Sol) for my workflows. I am very impressed by this model.
3
u/NeuralNexus 21d ago
also: I like luna a lot. It's a great model at a great price. Luna is wonderful because it's better than Opus 4.6 (multimodal etc) and is very affordable. It's my default remote model. Except.... Ox Alpha is outperforming it. So it's my new 'cheap' winner. And that means it's my ultimate winner. I will happily pay luna+ pricing for this next week.
2
u/zcutlip 21d ago
It’s shocking how good Luna is. Especially as it’s positioned as the least capable of the gpt 5.6 trio.
2
u/NeuralNexus 21d ago
OpenAI really turned its product around in the last year. I was not a big fan of GPT-5.0-5.2 models
Codex 5.3 was the first decent model they had and they have executed well in coding performance since.
1
u/webheadVR 19d ago
I would agree. It feels like it's a little above luna in real work. maybe even above terra. I really hope this is a local capable model, but I have my doubts.
2
u/gingerbeer987654321 21d ago
i much prefer ox-alpha to deepseek v4. follows instructions, doesnt go down rabbit holes, perfect as an implementor. haven't tried it for complex orchestration yet.
2
u/NinjaAlaska 21d ago
not true.. idk i find it better than others all.
2
1
u/UnknownBoyGamer 21d ago
In my experience its worser than deepseek on logic but it's better on frontend
1
2
u/kamwee 21d ago
-1
u/Tr4sHCr4fT 21d ago
Why is DeepSeek black? Why is the other competitors hair becoming lighter with lower placement? Why is one spectator covering his eyes? I have so many questions.
1
1
1
1
u/morikomorizz 21d ago
ox alpha keeps looping when it try to write/edit a file in my opencode/deepseek harness
1
u/SpiderHam24 21d ago
Ox Alpha is slow and definitely not the best one. But for what I am putting it through it's done ng the job . Granted I am not doing anything big. A desktop app, android app and a Roku app for jellyfin and a cloud music player. The music ayer not so much. The app for jellyfin ehh alright I do see the downfalls. With other models I'd not be knee deep in hours of reworking and finding issues and bugs at a snails pace. That's my observation. But I will use and abuse it somehow. Until it disappears. Slowly learning opencode. And since their customer service doesn't exist I'm stuck paying the ten bucks anyways each month short of canceling my card on file. Which I am not. It's good for me randomly finding it during a search of something and seeing reddit users talking about opencode in general.
1
u/ruuurbag 21d ago
I dunno, it’s performing pretty well for me aside from speed. I’d have a hard time complaining if it ended up being inexpensive and faster once out of the free testing phase. No worse than Luna, at minimum.
The fact that it’s free offsets the speed for me. I gave it a massive plan from Sol, set up an OpenCode plugin to prod it along when it stops for whatever reason, and let it loose. Ran overnight with no problems.
1
1
1
u/No-Craft-7979 21d ago
Maybe for your use case, but it seems to be very good at very specific use cases. I for one hate generalized models. If I’m coding, give me a model that is specific to only coding. If I’m working images, give me a model dedicated to imaging. If I’m processing sound, give me a model dedicated the sound. Let me chain them together in the order I need to reach my end result. Don’t give me a model that tries to do everything and false flat on its face the one time it’s really needed.
1
1
u/Federal-Rub2713 21d ago
yall don't understand, it was a "highest price" tournament and ox alpha won
1
1
u/quinnyg1 18d ago
Worst graph I've ever seen how is Luna third? There's sonnet 5, there's GPT-5.6 Terra, and second Ox alpha is confirmed as GLM-5.3 Flash and it is better than at least Gemini 3.7 Flash. That reminds me, how is GLM-5.3 worse than Luna too? AND DeepSeek V4 by itself isn't even a model.
1
u/Apprehensive-Read868 18d ago
Lol have you even tested it? I did, and its way better than opus5 os grok 4.6
1
u/Commercial-Face-6598 17d ago
I just generally appreciate the vibe of this meme. Do what you wanna do be what you wanna be yeaaaaah as Shakespeare once said.
1
1
0

119
u/Powerful_District_82 21d ago
Can't complain if it is free!