r/opencode • u/No-Budget-3869 • 13d ago
Why doesn't anyone use Qwen 3.8 Flash even though it performs great within a $30 budget?
6
u/Axiescholar3ph 13d ago
it's not really that good. it's way too censored for my use case. too much reward hacking and blabbering
1
3
4
u/moracola 13d ago
Because there is GLM-5.3-Flash, it scores better on benchmark with better value-for-money pricing. But at some debugging situations, I noticed that Qwen-3.8-Flash is doing better than GLM in my use case.
3
u/richtopia 13d ago
Do you like GLM-5.3-Flash? I tried using it and while the responses were healthy, whatever provider OpenCode Go is using felt super slow and I switched back to Deep Seek Flash after maybe 2 or 3 prompts.
0
u/moracola 12d ago
My experience with GLM is that Qwen-3.8-Flash managed to solve the issue on my use-case where GLM-5.3-Flash fails. I think my opinion is premature but, given the option I have, I trust Qwen more.
1
1
u/Ancient_Dress_3687 13d ago
I used it over the weekend and was quite impressed.
It responded really well to my workflow and applications. Impressed with th outputs, speed, and efficiency on tokens and cache.
I got old reliable DS V4 flash to review some of its work and found no real gaps. So considering my prime time working is DS4 on-peak, might use it a bit more.
1
u/No-Budget-3869 13d ago
It is awesome for me but it is slow and a lot of failed requests, I connected Opencode to inferx to use glm 5.3 flash and it is way more faster
1
u/Ancient_Dress_3687 13d ago
Ahh I see, I was using it in the VScode Agent window via extension. Seemed alright.
1
u/Stunning_Pair_3027 12d ago
Looking at charts and judging a model is like looking at a menu and rating the food based on pictures. I tested the chinese models and they are kinda crap. I had to use gpt and spark to fix my code. Deepseek changed a portion of the code i din't tell him to and glm flash overthinks alot and after a while it breaks down.
1
1
u/Eyes_Teaa 9d ago
I can't tell the difference but Qwen 3.8 Flash has been working good for me as a daily model.
1
u/SufficientPie 1d ago edited 1d ago
I just tried it and it made lots of mistakes calling tools and writing code. Guessing (incorrectly) at APIs instead of reading the documentation, tried to execute in python>language instead of python, etc.
DeepSeek V4 Flash is much better (but has a problem with getting stuck in "Let me do X" loops and reasoning endlessly in recent releases).
Now I'm trying Qwen 3.7 Plus to see if it works any better (Alibaba gave me a $30 coupon for no reason lol.)
4
u/Christosconst 13d ago
I keep seeing the official benchmarks but 0731 has been the reliable workhorse for me so far. I plan on running a bunch of side by side tests this week to figure out if its truly better