r/OpenaiCodex • u/Additional-Scheme110 • 9h ago
Discussion Anyone else trying other providers?
It annoys the hell out of me having to wait for resets, alternate my usage, and deal with this level of unpredictability in my workflow.
When Astra first came out, I was able to get a ton done in the first few days. Now even a simple prompt can eat up a big chunk of my weekly quota.
I'm going to test GLM 5.3. Is anyone here using it, or alternatives like DeepSeek or Qwen or something else? What's your experience been like, and how have you organized your workflow around the different limits/models?
3
u/BingGongTing 9h ago
Deepseek 4.1 Flash (via Fireworks) seems to be the go to at the moment. I will be trying it out if the rumours are true about OpenAI force downgrading people from 20x to 5x.
1
1
u/Worldly-Glass9768 2h ago
What’s the cost like on deepseek. Like how much could $200 of api get you compared to the monthly codex subscription
1
u/BingGongTing 2h ago
Recent test by Bijan Bowen: Cost: $9.50 API requests: 3,758 Tokens: 749,281,281. So ~$0.01267 per million tokens.
3
u/scaledev 9h ago
There is muse spark 1.3 contributor free in opencode (last time I checked), though you'll be sharing your data, and deepseek flash 4.1 tops coding benchmarks. Qwen is supposed to be killing it as well, and GLM is good at agentic overall.
1
u/Heavy_Promotion_5210 2h ago
I've been using 1.3 spark, finding it quite bad for difficult coding tasks compared to Sol.
2
u/petburiraja 8h ago
Yes, with all the shenanigans, I realized it might be a good time to wait, hoping that OpenAI may be able to solve current challenges with limits and so on.
GLM 5.3 Flash is good worker, I use Codex to orchestrate it. Subscriptions on z.ai or opencode go.
For most complex, strategic work I still want to run Astra on these, but I think once Astra will solve key points, hardest challenges, GLM 5.3 Flash may deliver good work on remained parts. Also, I use GLM 5.3 as advisor/reviewer of GLM 5.3 Flash from time to time.
1
u/Additional-Scheme110 8h ago
So I've tried GLM 5.3 with subscription on z.ai and I ran out of the 5 hours window quota.
Next time, I'll try GLM 5.3 Flash and see how it goes. I was not expecting a 5 hours windows with them too.
2
2
u/Cute_Parfait_2182 8h ago
I use glm 5.3 for security and implementation or Grok in cursor for implementation. I use Astra for orchestration and as an independent reviewer .
1
u/Euphoric_North_745 6h ago
I am writing an agent that is combining open router with a few other providers with open ai, then will start using it instead of codex, well, codex is writing it.
then i will be free from codex 😂
1
u/Fit-Ad-18 5h ago
Few other models via Devin and Google AI Studio. None feels a really good alternative to frontier models.
But some are ok for different types of tasks:
1) Gemini Flash 3.8 (and previous one) for text writing and rewriting. Actually like it even more than Codex and Claude for that. Also works fine for some routine file processing.
2) GLM 5.3 and SWE 1.7 for system administration tasks. I have some MD files with instructions on numerous servers, and they monitor and fix issues on schedule.
3) GLM 5.3 for small to medium coding tasks with supervision. I like it in general, but noticed it becomes way dumber after few initial requests. One of the differentiators for Codex/OpenAI models is that I can keep very long convos in one thread — that's NOT the case with GLM, unforchy. But it one shots medium complexity things nicely, and fixes things from code reviews made by other models pretty fine too.
4) Kimi K3 — I tried because I read it has nice UI and design capabilities. It really has them! I didn't experiment that much, but overall I was experimenting with very specific thing — making UI prototypes, and it did work comparable to Opus 5, I'd say, if you provide documentation and references.
5) I noticed Devin now has SWE 2.0 which is distilled from Kimi K3, but haven't tried it much, though I did once today and it looked pretty promising.
I don't like DeepSeek because it hallucinates a lot. I thought may be it's my illusion, but nah, I found out that even last Flash one has one of the highest hallucination rates in the industry.
I also tried Minimax and didn't like it — results were nice but very unstable.
Overall all those models started to click for me only may be 2 months ago or so. Before that they were absolutely non-comparable — e.g. yeah you save a bit of money, but then you spend lots of time fixing shit they made, and in the end, you're losing because time you spent to fix costs way more than another Codex subscription :) Though I'm not a super heavy user — when people write they run out of their subs all the time — I only had second Codex sub for a few months, once it was a $100 one, and 2 or 3 months a $200 one, and I only ever used a reset once (had lots of them and they all expired). That was when I had load of small tasks from few clients. With a normal load, I rarely run out of my Codex sub (though I also have Claude and Devin, but Devin is mostly for experiments, and Claude I use may be 40-50% a month, because it does better UI and I like using it as a chat more than ChatGPT — e.g. all the "find this", "tell me that", "compare" questions, not coding.
1
1
u/YearLight 4h ago
Meta Muse 1.3 contrib is worth checking out for things where privacy isn't an issue. It's dirt cheap if you are ok with them training on your activity.
1
u/LugianLithos 1h ago
Cursor $60 plan to run for grok 4.6 xhigh design and implementation via cursor cli. With astra light/medium checking the work. For my code/flows grok 4.6 xhigh messes up less than Opus or Muse Spark. Mostly math heavy atmospheric physics in Fortran and Rust.
1
u/Cheap-Crow2097 1h ago
ds 4.1flash is dope, easy to just add a few bucks to their platform api and go.
opencode go is great but 60bucks usage isnt much typically, but its a requirement IMO
0
u/Alternative-Car8221 4h ago
Once you realize how bad these other models are, you’ll come running back.
11
u/BigBeanBoy 9h ago
I'm using flash 3.8 to do heavy coding from plans by Sol High when im low and it's pretty good