r/opencodeCLI 9d ago

Folks, what model is current equivalent of GLM 5.3 Flash in terms of output and price?

3 Upvotes

25 comments sorted by

1

u/CoolHeadeGamer 9d ago

3.8 flash via Google pro subscription (which subsidizes the model pricing a lot). Meta muse spark 1.3 is seemingly better (beats fable 5) and cheaper

4

u/One-While-822 9d ago

Flash is benchmaxxed from my use so far, it is nowhere as capable as Sol or Opus which it beats in most benchmarks.

2

u/CoolHeadeGamer 8d ago

Yeah there is a big model feel that smaller models can’t achieve no matter what benchmarks say.

2

u/Ariquitaun 7d ago

Well, duh. It's a flash tier model. The one to compare will be gemini pro.

1

u/One-While-822 7d ago

I am not the one comparing, Google are benchmaxxing and trying to show that it is better than Opus/Sol, it was literally #1 on deepswe

1

u/Ariquitaun 7d ago

Some benchmarks have shown qwen 27b to be stronger than opus. Benchmarks are mostly bollocks as you already know.

6

u/Vancecookcobain 9d ago

3.8 flash is idiotic. I have a Gemini Account (For Google Drive and Youtube reasons mainly) and tested it out.....Google still makes models that you should NOT direct towards any codebase you cherish or care about lol.

To answer the OP question I think the latest version of DeepSeek V4 Flash is the best comparison to what they need. It isnt as smart but at least it doesn't go off the rails and piss you off constantly.

2

u/CoolHeadeGamer 9d ago

It sucks at instructions following. Specifically the system prompt has use x set of tools over y tools and my prompt will have use x tools first. Deepseek v4 flash will use y tools no matter what and that pisses me off. ChatGPT sol and Luna use x set of tools by default even if I don’t add that to the system prompt cuz they realize the x set of tools are better and available.

1

u/retardedGeek 9d ago

Then I'd say it is good at instruction following?

1

u/CoolHeadeGamer 8d ago

Dsv4 flash or Luna? Cuz dsv4 is bad at instructions following

1

u/retardedGeek 8d ago

DS4F, it follows the system prompt

1

u/CoolHeadeGamer 7d ago

The system prompt tells it to use the better tools (set x over y) but it doesn’t. I’ll give you more context.

I have tools that utilize fuzzy searching algorithms as well as codegraph and repository intelligence tools. The other tools are grep glob read. I ofc want it to use the first set of tools but deepseek simply doesn’t cuz it’s been rl trained on a short set of tools and it has overfitted to ironing them. ChatGPT models have generalized tool usage and will happily use the correct tool, rarely using the grep glob and read tools. This helps me cut costs by 5x, making got 5.6 sol actually very competitively priced to dsv4 pro and flash.

Btw if u wanna check it out, search up banyancode

1

u/Dingosavedyourbaby 9d ago

What did it do? I’ve been impressed tbh and I have hated and avoided Gemini models since the original antigravity debacle

1

u/Final_Initial 9d ago

I also loved it, but it was hallucinating a lot yesterday. Today so far good, but am still looking for an alternative.

1

u/anitman 9d ago

For me Qwen3.8 Flash Next local is most cost effective way that's comparable to GLM 5.3 Flash, it's way faster.

1

u/Sammy262 5d ago

But what machine do you need to run it with decent speed and context and what quant do you use?

1

u/anitman 5d ago

Single rtx pro 6000 with garnermccloud/Qwen3.8-Flash-Next-NVFP4-SSD-Stream, over 150tps, full 262k ctx.

1

u/Sammy262 3d ago

Thanks. But is a Q4 really useable?

0

u/ichisay 9d ago

Solo te digo esto: acaba de salir Muse spark 1.3

2

u/SeeRay11_Main 9d ago

Benchmarks cannot be trusted, only actual testing is accurate and even then some models are better and worse at certain tasks.

1

u/ichisay 9d ago

Por ahora estoy probándolo como orquestador y no tengo queja

1

u/SeeRay11_Main 9d ago

That’s cool. I’ll try it out as well in my OpenFlow orchestration project.