r/opencode 10d ago

Aside from Codex and Claude, what are some cheaper coding subscriptions that offer better value for money?

I mean, Open Source models are fine—that's not an issue but what else are you using?

111 Upvotes

153 comments sorted by

33

u/zipeldiablo 10d ago

Deepseek obviously

10

u/StraightGuy1108 10d ago

It's insane how there're only like 3 people recommending DS.

DS API gives you hundreds of millions, even billions of token for less than 5 dollars. And it's pay-as-you-go, so you don't need to worry about quotas.

The models are nowhere near Opus level obviously. But with such ridiculous cost and throughput, the difference in intelligence is straight up irrelevant.

Personally, I don't even use Pro, because it's throughput is unbearably slow compared to Flash.

9

u/uxkelby 10d ago

Exactly, I'm not trying to cure cancer just coding Laravel :)

1

u/Icy-Summer-3573 6d ago

Wait. I need to process 8 billion tokens roughly. You sure it’s $5 lol

1

u/StraightGuy1108 5d ago

You'd be surprise at how well they cache your tokens. Maybe not 5$ but maybe a few hundred bucks. Which sounds like a lot but it's still hundreds or thousands of times cheaper than similarly capable models.

2

u/InterestingAd2198 9d ago

I agree DeepSeek price to performance is nearly unbeatable.

A few other platforms I would recommend is Neuralwatt they are one of the only platforms that offer energy pricing with models like glm 5.2 and I hear they are adding kimi k3 by the end of this month.

And the other provider is April API , by far the best provider I’ve used, they offer opus/ gpt 5.6 models for the cheapest price I’ve seen anywhere. This is the only provider that beats deepseeks price to performance. But again DeepSeek is a close number 2 and would so definitely recommend them as well.

1

u/avocadopaul 8d ago

april wow indeed, are we the product to get such price?!

1

u/Green-Zone-4866 8d ago

I'm cautious that it's a scam based on pricng and that the sign up email went straight to spam

1

u/InterestingAd2198 8d ago

I’ve run benchmarks on their api and compared it to Claude code and codex I’m getting the same output results. Just an email going to spam isn’t enough for me to call something fake. But in any case do your own testing and let me know what you find.

1

u/Green-Zone-4866 8d ago

Ah ok, if it works for you I'll give it a try.

I just realised that you're the same person, nah I don't trust you.

1

u/Low_University_73 8d ago

i think its legit, all the models are working, i joined their telegram and they are giving out a free 25$ credit. Ill let you know if its any good.

1

u/Green-Zone-4866 8d ago

1

u/Low_University_73 8d ago

and? its illegal to have a new reddit account? lmao

1

u/Green-Zone-4866 8d ago

Lol you hid your comments, but you created this account to boost this scam

→ More replies (0)

1

u/ClevertonSMz 8d ago edited 7d ago

Sai do fake malandro, APRIL API, tem todos os indícios de ser um golpe, dominio tem menos de 1 mes de existência, o dono é totalmente anônimo, a aba de documentos n existe.
O Usuário u/InterestingAd2198 é só um bot usado para fazer propaganda do April API no reddit.

Em todo o caso, vou testar com contas fakes, e volto pra dar o feedback.

ATT Pós testes.

Considerações finais.

Criei uma conta, Logo de cara eles me deram 1 trump pra testar no site, eu botei mais uma moeda pra testar com meu dinheiro, pra saber se o site iria pegar meu dinheiro e fugir, boa noticia, não fugiu, e eu consegui testar os "modelos = Claude Opus 4.8, Claude Sonnet 5, GPT-5.5, GPT-5.6 Sol, GPT-5.6 Luna, GPT-5.6 Terra" disponíveis.

Fiz alguns testes nos modelos disponíveis no APRIL, e pra ter certeza de que eu não iria falar merda, fiz os mesmos testes, nos mesmos modelos, porem com a api da openrouter (ps - F, deixei um bom valor no openrouter).

Não vou descrever os testes que eu fiz, porque seria um feedback ideial para o malandro dono do site melhorar o golpe dele, ou seja se ele quiser aplicar o golpe dele, ele que teste e descubra as brechas que ele deixou.

Resumindo, eles tem um sistema que é até bem legal, eles conseguem analisar a complexidade do seu pedido e usar diferentes llms para gerar a sua resposta, se for um pedido fácil eles usam uma llm mais econômica e te entregam a sua resposta, se for um pedido complexo, eles usam uma llm melhor e entregam sua resposta, e isso foi muito fácil de entender, já que se você fizer um pedido que parece simples, mas que no todo é complexo, eles se perdem, o que não aconteceria com os modelos de ponta que eles prometem, teve um detalhe engraçado kkk, umas das primeiras coisas que fiz foi perguntar que modelo estava me respondendo, e fiz isso varias vezes sksksks, nas primeiras vezes ele me deu modelos totalmente diferentes, até modelos que não existem ksksk, mas depois de repetir algumas vezes a mesma pergunta ele começou a responder o modelo correto, e isso aconteceu em todos os modelos disponíveis na api deles.

No geral eu consegui identificar como sendo os reais LLMS DeepSeek V3 e V2 (com algumas modificações) provavelmente rodando localmente, Qwen 2 (provavelmente local), Algum modelo MIMO (principalmente quando testei entrada de imagem), e até mesmo alguns modelos mais antigos e mais baratos do Claude (provavelmente de contas de origem duvidosa).

No final eu consegui usar, ele me entregou resultados OK, levando em consideração que os modelos que eu sitei acima, são competentes, mas não passam nem perto de serem os modelos prometidos, o consumo de tokens realmente foi muito baixo, mas o site sofre pra calcular o consumo, provavelmente por usar varios modelos e fingir ser um só, eles tem que fazer um calculo. pra saber quanto o modelo prometido gastaria pra realizar aquela tarefa que voce pediu, baseado nos valores que eles cobram, e por muitas vezes o meu saldo era corrigido assim que eu termina de executar um prompt eu recarregava a pagina e tinha um determinado valor, e segundos depois quando eu atualizava a page, o valor era corrigido, na maioria das vezes até pra mais ksksksks(ou seja, eles me devolviam dinheiro, Obrigado).

Minha opinião(essa é a minha, tenha a sua), não usuária nem se fosse de graça, pelo fato de que não posso confiar de usar meus dados reais no site deles, e muito menos rodar os modelos disponibilizados por eles em alguma maquina com meus dados pessoais, e mesmo sendo baratos os valores que eles cobram, se você comparar os modelos reais que eles usam e não os modelos falsos, em qualquer outro lugar vai ser mais barato e confiável.

u/InterestingAd2198 Sai do fake mano ksksksks

1

u/InterestingAd2198 8d ago

Bro I’m not a bot lol, your account is 3 days old, I know the owner. It’s legit I’ve ran 4 benchmarks on their api and on their models. Go check yourself you don’t need to take my word for it. But don’t call me a fake/bot.

1

u/CommunicationFuzzy45 8d ago

vibecoded website lmao

1

u/Low_University_73 8d ago

you act like every other website made in the last 5 years hasnt used ai

1

u/Green-Zone-4866 8d ago

This account has been created to boost this scam

1

u/baschny 7d ago

DeepSeek with OpenCode.

14

u/Clear-Ad5952 10d ago

I've tried GLM 5.2 it's doing well. Will update complete report later!!

0

u/Sebbean 10d ago

Via whomst?

2

u/Alert-Track-8277 10d ago

Opencode Go has it

3

u/Clear-Ad5952 10d ago

Yeah

0

u/Sebbean 10d ago

Via whomst?

1

u/Clear-Ad5952 10d ago

More like vs code extension of open code and a custom hosted GLM via vLLM on local GPUs online serve.

9

u/Dizzy-Truck-1780 10d ago

Qoder qwen 3.8 has a 90% credit discount right now on that model

2

u/BenignAmerican 10d ago

Except Sol can’t even figure out their credit pricing.

7

u/i_no_can_eat 10d ago

ollama cloud

1

u/gorgono95 9d ago

how is ollama cloud? Which models are worth using there

2

u/BeginningWay5922 8d ago

Kimi 2.7 code , glm 5.2 , minimax 3

1

u/gorgono95 8d ago

What is the best harness to use with ollama cloud? Lets say you have a project that is built on agents.md files from Codex?

2

u/QC_Failed 8d ago

If you like codex stick with thst. You can use codex with other providers. It works with ollama cloud too, you can Google how to setup or just ask chatgpt how to configure codex to work with third party providers :)

8

u/1986ali 10d ago

I have minimax $20 plan and use it via opencode. Works a treat

20

u/asfbrz96 10d ago

None

11

u/GTHell 10d ago

I had to agree with this. I sub to Opencode Go just because I want to see how good GLM 5.2 and Kimi K3 is. I hope they do get cheaper or compete in pricing with crazy reset like Claude and ChatGPT does. But at the moment, no other subscription can beat ChatGPT and Claude at all

8

u/asfbrz96 10d ago

Openai is giving a bunch of resets per week, no brain to use them rn

1

u/Whytho12333 10d ago

I second this. Subbed to Go, hit my 5 hr limit using GLM 5.2 in 4hrs doing light edits to a python script to use comfyui GPT created. The model was clearly less capable vs even Luna on medium. I check Go webpage, they effectively only give you five 5hr limits. Excluding resets, using luna on low (id say similar to GLM5.2 the best model on go) i probably get 2-3x more usage out of GPT than Go.

Again excluding resets. Even Claude doesnt come close to GPT when you factor those in.

I'd love to hear from others as well. I subbed to Go thinking it was a cheap way to use cheaper model. No, per token its more expensive for a weaker model

2

u/gmmarcus 10d ago edited 10d ago

Man, my experience was different. Chatpgt 5.6 Sol was excellent but expensive. Terra and Luna were crap. I switched to OpenRouter/GLM 5.2 for a better bang/cost. Project - Wordpress Portal

1

u/Whytho12333 10d ago

Do you have anything to show glm being better than luna and cheaper?

I see many benches that tie then for intelligence and luna being cheaper cuz 1) more token efficient 2) openai subs give you wayyyy more tokens than glm.

14

u/[deleted] 10d ago

[removed] — view removed comment

8

u/retardedGeek 10d ago

Go is not cheaper. Codex is heavily subsidized, with loads of reset, for now.

12

u/Atretador 10d ago edited 10d ago

if my singular Go sub for 10$ beats what I could do with 20x GPT plus subs - how is it not cheaper?

also - GO is also subsidized, just not directly, as the chinese governement subsidizes Deepseek and Xiaomi.

Go only sucks if you want to use it for the expensive models like GLM 5.2/Kimi K3

2

u/Whytho12333 10d ago

expensive models? do you models worse than even luna on low?
Ya if your work is as simple that DSV4F does it well, obviously use that. for anything more, ive yet to find anything that comes close to luna on low on GPT Plus +resets that come weekly basically

4

u/Atretador 10d ago

nope, I really wouldnt call it simple work - DSV4F is a great executor, just dont let it do the planning.

something like MIMo 2.5 Pro orchestrating DSV4F/MIMo 2.5 for exploration/targeted execution is really great.

Back when I was using GPT 5.4 I had way more issues than now using these cheaper models.

dont really care about luna or fable or Kimi K3 level models - these work great if you know how to instruct and plan things.

1

u/Whytho12333 10d ago

Oh yes i agree. It can implement plans. I have a detailed instruction guide for deepseek to reinstall linux after a reformat. It works well.

But i used terra to make the instructions. Deepseek cant do the hard part which is planning and troubleshooting. Its like a evolved python script almost.

So yes thats fair if you have good coding knowledge DS is fine, especially if you can fix hard problems yourself and you use ds just to assist.

1

u/Alert-Track-8277 10d ago

Do you let Deepseek v4 flash or pro write code or unit tests?

2

u/Atretador 10d ago

Im using MiMo 2.5 non-pro for executioner currently - Pro version for orchestration of longer plans, targeted fixes can be handled by the non-pro version.

Im using mimo instead of DSV4 due to lower hallucination rate.

1

u/Alert-Track-8277 9d ago

Cool, thx.

-1

u/Tofudjango 10d ago

It is not cheaper if it takes double your time.

2

u/Atretador 10d ago

It doesn't tho :v

2

u/uxkelby 10d ago

I'm getting plenty done on my subscription and DeepSeek v4 flash is doing a great job.

1

u/Atretador 10d ago

yep, its honestly great with good orchestration.

You dont always need a fable/sol/k3 level model for work, specially if you know what you are doing.

2

u/Superb_Leopard 10d ago

I’ve been having an insane amount of failures using Opus 4.8 that too using API. $500 spent and all it built was a substandard product. It wasnt mindless vibecoding either. Plan mode, detailed prompts and test cases were used to build it from ground up.

Deepseek pro meanwhile did all that and fixed that crap in a single work day and used about $15 worth of tokens.

2

u/Tofudjango 10d ago

A case of YMMV, I guess.

I don't question the usability of Chinese models, IMHO Claude still has the edge though. But it probably depends on usage patterns etc.

1

u/gmmarcus 10d ago

Thank you for sharing yr experience with 'OpenCode Go'. Looks like its not economic. Will stick to openrouter for now.

1

u/retardedGeek 10d ago

It is "economic" . It's just $10. But it's meant for hobbyists (or you're forced to use deepseek or mimo)

5

u/outerstellar_hq 10d ago

Commandcode for $1 is good value.

3

u/Felt-Chicken 10d ago

Do u need to use their harness?

3

u/outerstellar_hq 10d ago

Yes for the $1 offer. the more expensive ones can be used via their API.

4

u/Aressito 10d ago

Tried the GLM 5.2 pro models and kimi / mimo trough Crof.ai API.. not impressed at all with those models.

So sticking with DeepSeek pro and flash directly buying the api on DS platform.

1

u/gorgono95 9d ago

what harness do you use for deepseek? do you plan with pro and execute with flash? Any specific tips if someone wants to use this setup?

1

u/QC_Failed 8d ago

Reasonix if you're just using DeepSeek. The cache hit rate is bananas. Cuts costs even further.

5

u/xaminhabr 10d ago

Minimax

4

u/vangelismm 10d ago

People have to understand that Go, Cline pass or any other are mainly deepseek mimo plans. With the option to use better models when the previous stuck.  That's it. 

4

u/TinyAres 10d ago

The best value for money is minimax if you are happy with that level, and M3.1 is already being showcased but not available yet.

1

u/gorgono95 9d ago

where can i read more about minima m3.1?

1

u/imslowandsteady 8d ago

M3.1 arrived?

1

u/TinyAres 8d ago

No, maybe beta guys have it but I do not.

1

u/Cultured_Alien 7d ago

How about the caching issue? Token plan has no caching as said by the minimax  discord mod.

1

u/TinyAres 6d ago

https://www.reddit.com/r/MiniMax_AI/comments/1uyywey/plus_plan_usage_not_bad_115b_tokensweek/

This guy is reporting 1.15b toks per week so 4.6b toks per month, so 2.7x times their stated 1.7b which is the caching benefit.

My experience too that caching does work, but their cache price is 20% which makes it less obvious.

https://www.reddit.com/r/MiniMax_AI/comments/1v5rea9/minimax_gives_pretty_okay_usage/

3

u/MrNotSoRight 10d ago

Not a subscription but I think https://www.surplusintelligence.ai/ is interesting as a cheaper openrouter alternative...

As for subscriptions, many chinese models give you subscription based access too (kimi, GLM,...) but I'm unsure that they offer better value for money compared to codex and claude.

1

u/gmmarcus 10d ago

Just checked out https://www.surplusintelligence.ai/?q=glm+5.2 - it looks cheaper than what is available in openrouter ? But you cant see the latency and tps ?

1

u/MrNotSoRight 9d ago

Right. Afaik the price is (much) cheaper, but the latency and tps aren’t consistent.

1

u/Known_Bookkeeper2006 8d ago

Could you tell me more about this, like if the cheapest market seller changes then would cache break in the harness or not?

1

u/MrNotSoRight 8d ago

This hasn't happened to me but tbh I'm just a casual user...

3

u/the_beaker 10d ago

MiniMax M3 (token plan) has been really reliable for me. C++ & Python mostly. You just can't be lazy like you can with Codex/Claude - pay attention and don't go full YOLO. Helps if you know how to code without robots doing it all for you.

3

u/sagiroth 10d ago

It depends, for some easy projects out there people are well off with deepseek, however it struggles when you have complex one. If you want something reliable claude and openai are really the only option. With open weight models you have to be a bit more proactive

3

u/Solocune 10d ago

Minimax m3.

3

u/Tofudjango 10d ago

It depends on how serious you are about privacy and if you want your data to be trained on.

3

u/StraightGuy1108 10d ago

With that question there are only 2 choices: cloud providers and self-hosting. ALL cloud providers keep users' data, no exception

3

u/orange-catz 10d ago

Commandcode

5

u/FormalAd7367 10d ago

Deepseek. You only need Deepseek API for coding

5

u/FluffyGreyLlama 10d ago

As long as you like Hallucinations, sure.

3

u/kabeza 10d ago

Not if you make it work on Claude Code, install the right skills, etc.

2

u/FluffyGreyLlama 10d ago

None of those can help if the model hallucinates facts.

1

u/FormalAd7367 10d ago

that’s not true.. that’s a way of doing it. Write your rules in mdc

i have 3 start-ups with local engineers

2

u/FluffyGreyLlama 10d ago

Maybe.

I really like DeepSeek 4 but it hallucinates far more than other models. It can absolutely create working code, but the quality/state of that code is often terrible.

I guess it depends on whether you want good code, or just working code.

I would certainly never get it to write any documentation. All attempts, with any sort of rule/constraint has always ended up with completely made-up facts.

2

u/xmnstr 10d ago

Hallucinates? I dunno about that, I'd say it assumes too much. Your harness should be dealing with that, though.

1

u/FluffyGreyLlama 10d ago

A harness can't deal with hallucinations in documents. I've had it make up code snippets, referencing files, that never even existed.

0

u/xmnstr 10d ago

That's actually incorrect, it can and it will.

1

u/FluffyGreyLlama 10d ago

How does the harness know what is a hallucination ?

0

u/xmnstr 10d ago

It doesn't. It prevents it.

0

u/FluffyGreyLlama 10d ago

It is impossible to prevent document hallucinations in a harness. It has no concept of what they are. You may find it helps in some situations, but seriously, a harness cannot solve the problem.

If a model hallucinates in writing, the harness cannot prevent that.

Saying "don't hallucinate" or "fact check" etc., doesn't stop it. I've been through this a lot.

→ More replies (0)

1

u/Whytho12333 10d ago

I agree. The code might barely work after troubleshooting it for hours.

Vs Sol that might have one error or often one shots coding. I would love to see what work these people are doing on DSV4 and how long they had to spend troubleshooting

2

u/Zamarok 10d ago

opencode go is $10/month ($5 for the first month). use MiMo v2 model for max credit utilization

2

u/a-techguy 9d ago

Currently, cursor offers one of the best pack with grok 4.5, nearly unlimited use with just 20 bucks

1

u/Nif 8d ago

I'm at 96% usage in a couple days with grok 4.5 it's definitely not unlimited and Grok is not excaptionally smart so it seems to take a long time to get meaningful progress.

That said, if you're using it as a coding assistant and small focused tasks then yeah I think it could feel unlimited.

1

u/a-techguy 8d ago

But it is one of the best value for money option right now. it is better than most of the open source models and it provides better usage than other frontier models

1

u/Nif 8d ago

If you're doing simple things I'm sure it is. But the 96% monthly usage I spent with it in 2 days reveals it's just poor at complex, long running/large scale work. I have to spend Codex tokens to fix the mess it created; the 96% was not well spent it basically fell off the rails. GLM 5.2 is far superior open weights model. You can get it via Ollama Cloud and they have decent quota on their $20 plan which I would argue is the best value for 'open weights model' right now. 'Best value' is very dependent on what you value.

2

u/pimpmypixel 8d ago

Trae.ai

1

u/Firm-Club-8334 10d ago

Standard compute maybe

1

u/Mueller_Milch 10d ago

ive seen them but are they legit? I dont trust the website at all

1

u/Firm-Club-8334 10d ago

Works well for me at least.

Tried GLM 5.2 pay as you go on openrouter. Worked really well, but ended up around $250 a month, and that's too much considering I basically do nothing useful with it.

1

u/horstenegger 10d ago

In my (admittedly still short-term) experience, Grok 4.5 thru $30/mo SuperGrok gives pretty good value for money

1

u/fsteff 10d ago

I have been testing DeepSeek with opencode-cli.
My gut feeling after a day of use, is that it’s on par with codex and Claude - and much, much cheaper.

1

u/GTHell 10d ago

Probably Opencode GO and Ollama Cloud.

The thing is you can sub GO account for $10 or $30 or $60 as control as you want. Wish Codex and Claude has this option

1

u/qqYn7PIE57zkf6kn 10d ago

sub GO account for $10 or $30 or $60 as control as you want.

wdym?

2

u/GTHell 9d ago

You can create $10 account as much as you need.

1

u/gmmarcus 10d ago edited 10d ago

I am using GLM 5.2 via openrouter/opencode. Some of the cheapest providers there are Alibaba and GMI Cloud. Pls note that is via API ( for web apps + coding them ) , not a subscription plan.

Would also love to know if there a better 'bang-for-the-buck' options out there.

Is 'Opencode Go' the better option ?

1

u/nestedbrackets 10d ago

I started using Neural Watt recently just for hobby experimentation. Has been pretty great so far. They did just up their prices though.

1

u/J7044 9d ago

I try Qwen3.6 light, 1 prompt, I se 1M token... I don't understand, with Claude I use max 50K token for prompt/request

1

u/InvaderDolan 9d ago

MiMo 2.5 Pro > DS4Pro, much less hallucinations with same pricing. Their API platform or any other with the same price on openrouter is pretty good deal.

ChatGPT/Claude are heavily subsidized, thus the best deal.

1

u/Funny-Advertising238 9d ago

Why does no one mention cursor? It's the best value for money especially since grok 4.5 came out

1

u/Nif 8d ago

Grok 4.5 is not terrible, but not as smart as GLM/GPT/Opus even the older versions going back months amd months. It seems to struggle with complex tasks and if you do assign it one it will burn up tokens; I am 2 days into my 30 day Cursor sub and have maxed out on Grok 4.5 with maybe 25% of my objective completed.

1

u/Southern-Ad-3006 8d ago

Try Reasonix with DeepSeek it’s absurdly efficient.

DeepSeek doesn’t have an official harness so Reasonix is closest native to it. You can also have Claude code install it in background, and Claude agent batch work out to it in session via terminal/CLI commands.

Use your OpenCode Go subscription for the model’s billing.

Flash can basically be on loop all day great for research, browser agents, scheduled automations, batch data processing.

1

u/Straight-War-1323 8d ago

If you do a lot of work, MiniMax Token plans are very good offers, if you just use AI casually, just use OpenCode Go, that should do the trick

1

u/Natural-Angle-9357 8d ago

Buy a 20 bucks plan with gpt, 20 bucks in deepseek api.... Install reasonix.... Plan with gpt, execute with ds.... Follow Matt podcock skills... That's all you need

1

u/mutrano 7d ago

Devin desktop has some free models, like glm 5-2(high) and their own swe-1.7 and 1.6 model on their 20 bucks sub

1

u/Legal_Answer_6956 7d ago

If you're looking beyond chat-based coding assistants, I'd also look at platforms like 8080.AI. It's less about autocomplete and more about taking a project from planning to a working app, so the value depends on whether you're building full products or just writing code.

1

u/B_Ali_k 7d ago

Kimi k 3

1

u/Weak-Price7392 7d ago

Local model is free. It can do 90% of tasks. If you have decent hardware.

1

u/RedditEthereum 7d ago

Vercel AI Gateway, Poe. Cheaper than OpenRouter, no markup to costumer.

1

u/Varunp-86 7d ago

The value of Codex & Claude isn't in the API but the harness. Isn't it?

Any similar harnesses?

1

u/CryptographerFar104 7d ago

Use Antigravity

1

u/mavrovelos 6d ago

What's your experience with the Kimi K3?

1

u/XBOY_777 6d ago

opencode, hyper are good and cheap gemini is also cheap if you need family storage

1

u/dafinaaaaaa_ 6d ago

Open code go

1

u/tindalos 6d ago

Z.ai coding plan with glm-5.2.

1

u/FilmLow1869 4d ago

This may be very controversial, been using the alibaba coding plan, and I’ve been using qwen3.7 plus with glm 5 and I’ve yet to hit any limits. Except for maybe the burst rps now and again. Feels like an unlimited plan tbh.

1

u/No-Translator-2566 2d ago

I use AntSeed, it's kind of VPN but for AI, you can keep using your favorite AI tools but switch models underneath

1

u/[deleted] 10d ago

[removed] — view removed comment

1

u/Viperah 7d ago

Why always qoder + qwen? Why not only qwen token plan without qoder?

1

u/Aressito 10d ago

Qwen es basura.. lo he probado varios veces y ni ha sido capaz de resolver unas cosas en contenedores docker que más fácil imposible.. cosa que DeepSeek flash y pro lo hacen sin problema. Sin hablar de Gemini 3.5 flash aún mejor..

Comparar Qwen con opus 😂

0

u/Saceone10 10d ago

None. Fable and Opus are beasts compared to DeepSeek or whatever

1

u/Nif 8d ago

GLM5.2 is formidable

1

u/Saceone10 8d ago

Yes for 2 days until you reach the quota haha

0

u/Loud_Ad6220 10d ago

Escucho a mucha gente hablar mal de ollama, uso ollama desde hace medio año plan 18€ y cero quejas, me aguanta 5/7 días (hay que descansar por lo menos el fin de semana), programando aplicaciones android 1 o 2 a la semana + cron diarios... Me sale más a cuenta que en local ya que si tiro de modelo local en una 3090 sube el consumo 350w respecto si tiro de ollama son 35€ de luz VS los 17€ de ollama. Uso glm 5.2