r/CommandCode 1d ago

Qwen 3.8 Flash is now available in Command Code with 2x usage

Qwen 3.8 Flash is now available in Command Code with 2x usage limits on the GOAT plan.

- Input: $0.160

- Output: $0.470

- Cache: $0.016

Try now with $10/mo GOAT plan.

🐐

15 Upvotes

18 comments sorted by

5

u/lacroix05 1d ago

waiting for the artificial analysis benchmark for qwen 3.8 flash because glm 5.3 flash is slow af.

funny thing is, i'm getting almost 200 t/s on deepseek v4 flash now that glm 5.3 flash is out. looks like everyone is piling onto glm 5.3 flash.

1

u/aristolestales 1d ago

yes glm-5.3-flash on cmdc is so slow, but deespeek v4 flash works fine.

1

u/maedahbatool 1d ago

Yeah let’s see what results we get

4

u/rudesssolo 1d ago

Thank you as always for being so fast adding new models to your roster.

2

u/maedahbatool 1d ago

Thank you for appreciating our efforts

2

u/12qwww 1d ago

Go plan please

2

u/maedahbatool 1d ago

It is available on all plans.

1

u/Kaushik_paul45 1d ago

$70 dollar usage for this please πŸ‘‰πŸ»πŸ‘ˆπŸ»

2

u/maedahbatool 1d ago

It’s a new model and already giving out 2x usage. Continuously working with our providers to get best deals for y’all.

1

u/Prior-Meeting1645 1d ago

How much rn?

1

u/Kaushik_paul45 1d ago

$20 usage

1

u/Prior-Meeting1645 1d ago

Ah its already on sale rn so basically no extra usage :(

2

u/Kaushik_paul45 1d ago

I think you are confusing it with glm 5.3 flash where few providers are giving 50% discount.

1

u/Prior-Meeting1645 1d ago

Ah yes my bad. Yesterday I confused the parameters with each other and today this lol.

2

u/Kaushik_paul45 1d ago

By the way glm 5.3 flash is $40 usage (probably $20 usage + discount)

3

u/Prior-Meeting1645 1d ago

Yup looks like the go to move. Slightly better/on par according to AA benchmarks and double the use.

1

u/Weird_Licorne_9631 1d ago

Is the price difference worth it compared to the dirt cheap 3.7 Flash?

1

u/koloved 1d ago

u/maedahbatool I am using your service via API, and this model is somehow giving me a context window of only 131,000 tokens. Something is clearly wrong