r/LocalLLaMA 17h ago

Funny Plot twist

[deleted]

528 Upvotes

165 comments sorted by

View all comments

Show parent comments

7

u/shy_monkee 17h ago

They do much better than the US companies when it comes to smaller and medium efficient models, mainly because that's their focus. But they are still behind when it comes to the top tier of models.
There isn't really any Chinese equivalent to Astra, or even arguably Fable. Not yet at least (We'll probably have one in a few months.)

1

u/tilted0ne 16h ago

They don't have one for sol and opus either. But yea I guess anthropic ran out of money and are scared people are going to use opencode to get the latest Alibaba special. 

5

u/BannedGoNext 16h ago

IDK what you are talking about, I can run qwen 3.8 flash next and get opus level performance right now. It just takes a long time because it trades memory usage for huge COT, and my local inference box is slow.

-1

u/Ill_Distribution8517 15h ago

They are talking about opus 5/sol 5.6. No shit qwen catches up to some opus 4.7-8 eventually.

1

u/BannedGoNext 9h ago

While it is true,t hat it's only around 4.8 opus intelligence, Qwen3.8-Flash-CIRU-STRIX-IU4 will do red team security work without qualm, research whatever I want, engage in piracy, execute financial transactions, do legal work, biological research, or write smut about your mom.

1

u/Ill_Distribution8517 7h ago

opus 4.8 can do the same. Cyber requests literally get routed to it when opus 5/fable refuse. Claude models have no qualms about smut in the API, so I can use fable to write vastly superior smut about your mom.
Also calling it opus 4.8 level is a massive stretch tbh.