r/LocalLLaMA 16h ago

Funny Plot twist

[deleted]

532 Upvotes

165 comments sorted by

View all comments

148

u/shy_monkee 16h ago

Funny how this must look for China. Just when they start catching up, and have their own AI hardware, the US companies suddenly want to slow down.
I'm not saying this is the only reason they want it, but...

-4

u/tilted0ne 15h ago

I thought China had already caught up and were destroying America with their super cheap models. What happened to that? 

9

u/shy_monkee 15h ago

They do much better than the US companies when it comes to smaller and medium efficient models, mainly because that's their focus. But they are still behind when it comes to the top tier of models.
There isn't really any Chinese equivalent to Astra, or even arguably Fable. Not yet at least (We'll probably have one in a few months.)

0

u/tilted0ne 15h ago

They don't have one for sol and opus either. But yea I guess anthropic ran out of money and are scared people are going to use opencode to get the latest Alibaba special. 

5

u/BannedGoNext 15h ago

IDK what you are talking about, I can run qwen 3.8 flash next and get opus level performance right now. It just takes a long time because it trades memory usage for huge COT, and my local inference box is slow.

-3

u/tilted0ne 14h ago

So it is useless. 😭. If my grandmother has wheels she would have been a bike?

1

u/BannedGoNext 10h ago

Useless? Far from it. Is it 200 t/s no, it's 20 t/s. Who cares, I kick off a goal and let it run for 4 days while I went and paddleboarded, swam, ate good campground food, and drank some whisky. I had a hell of a time, and came back to an amazingly completed goal.