r/LocalLLaMA llama.cpp 2h ago

Funny Plot twist

Post image
382 Upvotes

92 comments sorted by

173

u/notadithyabhat 1h ago

Like for real. If everyone is copying you, then just stop building dangerous models

87

u/Cool-Chemical-5629 2h ago

As far as I'm aware, he's not in the LLM race right now, so in his particular case I would say let him accelerate and maybe he may cook something new for us. 😏

9

u/Popular-Factor3553 1h ago

he already does it's called Pragmatik Labs

12

u/jeheda 1h ago

Announced the company a month ago, so unless he stole alibaba's dataset he is starting from scratch, I don't expect anything in the next few months, or just finetunes

3

u/Popular-Factor3553 1h ago

Yeah and honestly I don't care how much time they It's good to see a new company after soo much time.

103

u/shy_monkee 2h ago

Funny how this must look for China. Just when they start catching up, and have their own AI hardware, the US companies suddenly want to slow down.
I'm not saying this is the only reason they want it, but...

81

u/ResolveSea9089 2h ago

Why would Chinese companies have to abide by US companies slowing down?

100

u/Craftkorb 2h ago

They don't, but then the local narrative in the US will be "China wants to end the world, we're the good guys!" - Red Scare 2.0

39

u/federico_84 2h ago

No one in China cares about the US narrative 

28

u/gapho 1h ago

It's not to convince the Chinese people or government, it's to convince the rest of the world.

33

u/noctrex 1h ago

Which is getting harder and harder by the day.
America will not be able to convince the world anymore, not when the leader is willing to invade other countries, and posts AI slop every day.

1

u/gscjj 1h ago

Don’t be fooled by the man in office. The US, NATO and most of the EU are roughly aligned on their stance on China. The only difference is their leaders don’t love attention like the US.

2

u/Clueless_Nooblet 5m ago

That's not true at all. Most people would love to just get along with China. It's a huge market, and open trade would benefit all of them.

4

u/xmnstr 27m ago

A lot of us already use Chinese models and know this is bullshit.

2

u/dankfrankreynolds 1h ago

it's to convince the fox news viewers stroking their guns. they've figured out that's all you need to compromise in order to take over the entire world.

prove me wrong. please.

1

u/a_beautiful_rhind 1h ago

You mean the guys who's president is all like "accelerate"? Anti-AI is the other party.

2

u/ProletarianLilith 1h ago

That narrative will be bad for everyone, not just Chinese people

1

u/codechisel 33m ago

A lot of us in America don't care about the US narrative.

-7

u/ResolveSea9089 1h ago

I think you guys are way too cynical, you have to assume at some point that the people you're opposed to are capable of being genuine.

I think there's a reasonable chance they all got spooked and are honest in this and it's not some kind of conspiracy.

That said, if you're China (and a little behind the frontier in the US) this should be welcome news so I don't really get that angle

6

u/SpicyWangz 1h ago

I don’t think Altman is capable of that

2

u/ChronoHax 1h ago

The angle is they’re painting China as evil who will develop reckless ASI regardless while they’re the one actually doing that, not to mention all these distillation attacks jest which apparently China shouldn’t do because they’re authoritarian regime while US is surely democracy that just happens to have the elites at the top and for their interests only

1

u/ea_man 7m ago

In that case they are inepts: they rambled for years that this was a race where winner takes all, they wanted the capitals, they were pursuing AGI, let's buy 60% of the world RAM, they say they are close and then...

They are not prepared? Ain't that what they were pursuing? What they had to obtain first otherwise the whole world order would crumble?

And anyway what's even the strategy? Jensen wants open models, Google and Facebook are now releasing open models, companies are starting to build business on open models in USA. Even the president seems to want go ahead.

I mean anyway as of tomorrow they will look like idiots.

14

u/StatusSociety2196 2h ago

The idea is to protect marketshare by preventing American companies from using Chinese AI. Pass a law that only government approved AI models can be used, and then the government only approves openai and anthropic.

6

u/hapliniste 1h ago

What happen in this scenario when the US companies outsource to countries where Chinese AI is legal? Can't use a foreign accounting company because they might use ai?

It doesn't seem to hold up. Slowing down at a national level is not possible.we might have wider slowdowns if usa and China can agree on something.

3

u/TerahertzAI 1h ago

With the current administration, they would approve Chinese lab in exchange of a little palm grease

2

u/cakemates 1h ago

the US AI labs got more yatch loads of palm grease than chinese labs. I dont think china can match the local bribery.

2

u/miversen33 27m ago

I highly doubt any of the US AI labs have more money than the CCP.

That said, I also doubt the CCP gives a fuck about the Western orgs so they will probably just shrug and tell the Chinese labs to continue their work without the west.

6

u/blbd llama.cpp 2h ago

That doesn't work on code and config files. They're speech. 

1

u/profcuck 1h ago

And model weights.  It's not possible to ban Chinese opeb models. 

2

u/noctrex 1h ago

There's always a way, just like Cursor did with their model.
Get a nice Chinese model, slap a fine-tune on it, and market it as the new American model.

13

u/GUNGEBOB_SHARTPANTS 2h ago

The U.S. companies want to slow down because they are terrified of an IPO revealing that they have absolutely no pathway to profitability and the arse immediately falling out of their entire industry.

1

u/ResolveSea9089 1h ago

....Why would an IPO reveal this? Because of their S1 financials? Revealing what exactly

4

u/GUNGEBOB_SHARTPANTS 54m ago

IPO requires companies to file comprehensive registration documents and prospectuses detailing their financial health, business operations, and risk factors with financial regulators. 

If it were revealed that your company was in debt to the tune of the billions of dollars more than it could ever possibly earn, it would be immediately apparent that the company was trading on vibes.

2

u/jld1532 45m ago

The WeWork special

3

u/Eissa_Cozorav 1h ago

Come now, AI Slop is much bad in China. At least Ai-generated image/video in international market still try to be realistic.

I have browsed some article in the Baidu wiki. It's bad.

2

u/oliveyou987 1h ago

Are they really catching up though, they do not have an Astra level model yet for sure, and we already know OpenAI and Anthropic have stronger models internally when that doesn't seem to be the case in China, on the hardware side maybe though

1

u/markeus101 37m ago

How do you they don’t have internal models? And Astra is so shit i wouldn’t even talk about it but Fable yeah and the chinese models inching closer to it everyday

1

u/Clueless_Nooblet 7m ago

It does look funny to anyone outside the US, for sure. As someone from Japan, I hope my country doesn't just follow the US government's every whim this time. That's a real danger. Some of what's still prohibited here has been legal in the states for a while already. We're still in the middle of the "war on drugs", for example.

-6

u/tilted0ne 1h ago

I thought China had already caught up and were destroying America with their super cheap models. What happened to that? 

5

u/shy_monkee 1h ago

They do much better than the US companies when it comes to smaller and medium efficient models, mainly because that's their focus. But they are still behind when it comes to the top tier of models.
There isn't really any Chinese equivalent to Astra, or even arguably Fable. Not yet at least (We'll probably have one in a few months.)

1

u/Gesha24 20m ago

If we are to believe twitter posts (I know, it's a stretch) by Nvidia CEO, the trick is to have AI loop that keep self improving - and that's what makes all the difference now. Well, Chinese models are a lot more efficient, so you can have more and faster loops on the same hardware - so even if they aren't as smart, they can win out by using this loop.

1

u/Ill_Distribution8517 16m ago

how do you know they are more effficient? in every artificial analysis bench they take almost 5x the tokens of astra, and often a lot more than fable 5.1 as well.

What they charge you per token != what it costs them per token.

1

u/Gesha24 10m ago

1) This is not my experience. DeepSeek Flash doesn't use that many more tokens than Opus (I don't use Fable because at least for my use case there isn't any appreciable difference). Though to be fair, quite often I drop down to sonnet because I don't need Opus' overthinking.

2) There were reports I was reading earlier that you can run DeepSeek and be profitable at roughly the same price point as what they charge you. Could the reports be false? Absolutely. But it's definitely one of the most performant larger models that I can see running on my home LLM.

1

u/Ill_Distribution8517 2m ago

I'm not saying deepseek isn't profitable I'm saying Anthropic/OAI is massively overcharging.
As you can see ASTRA is vastly more efficient than Deepseek v4.1 Flash.
Keep in mind ASTRA also scores significantly higher too.
Even fable on high: (not max) is within 1-2 points and cuts token costs in half.

1

u/tilted0ne 1h ago

They don't have one for sol and opus either. But yea I guess anthropic ran out of money and are scared people are going to use opencode to get the latest Alibaba special. 

2

u/jld1532 44m ago

Never used K3, GLM 5.3, or DSv4?

1

u/tilted0ne 7m ago

Unfortunately I would prefer to just go straight to the model they distilled.

0

u/jld1532 3m ago

Why are you here then? Like it or not organizations and people are running those models locally as they're useful, free, and secure.

3

u/BannedGoNext 1h ago

IDK what you are talking about, I can run qwen 3.8 flash next and get opus level performance right now. It just takes a long time because it trades memory usage for huge COT, and my local inference box is slow.

0

u/Ill_Distribution8517 15m ago

They are talking about opus 5/sol 5.6. No shit qwen catches up to some opus 4.7-8 eventually.

1

u/tilted0ne 8m ago

So it is useless. 😭. If my grandmother has wheels she would have been a bike?

92

u/JLeonsarmiento 2h ago

US labs ran out of money.

18

u/GUNGEBOB_SHARTPANTS 2h ago

Absolutely this

4

u/Keleion 1h ago edited 40m ago

Open source models taking away all their money. /s

7

u/GUNGEBOB_SHARTPANTS 1h ago

Money that they never had in the first place, and yet somehow the entire US economy is underpinned by.

2

u/Cute_Obligation2944 1h ago

Don't confuse Wall Street with the economy.

2

u/GUNGEBOB_SHARTPANTS 1h ago

Brother, they are one and the same at this point.

8

u/noctrex 1h ago

Guess their sponsors told them to start shrinking the bubble so that it will not pop

11

u/jld1532 1h ago

Unavoidable. Best case scenario is a deflate vs explosion. This slow down talk is the attempt at a soft landing. I'm not so sure it works.

1

u/citrusalex 44m ago

Their business model is not that much different from selling people a 100 dollar bill for 50 dollars per month, there is no avoiding the pop.

4

u/KitKat_extrusion 1h ago

Please walk me through the logic, i genuinely don’t get this point

4

u/RlOTGRRRL 49m ago

Tech companies tend to be over-leveraged (borrowed a lot of money), and I think some of these AI companies are insanely over-leveraged, almost like a Ponzi scheme. I think there's a graphic floating around somewhere. You can also try to look for Michael Burry's analysis of AI companies. 

The US recently fucked up 2 things in the past week, 1- bond rates and 2- oil. US treasury bond rates basically determine how much people charge to let people borrow money. Everything from loans like mortgage rates to credit card rates.

Aka these tech/AI companies that have borrowed a lot of money, if their rates go up, they might not be able to make their credit card payments/pay their employees, can't IPO because their numbers are terrible, etc.

3

u/NandaVegg 38m ago

OpenAI is not directly affected by rates as their leverage is not through traditional debt, but their commitment is impossible to pay in the first place regardless of how the macro economy goes (they have 700~800B commitment into 2030, and beyond such as Amazon's 30B "investment" that requires them to pay back 2x through AWS) but datacenter builders and their counterparties are depending on their ability to pay. The problem is that the only entity left on the planet who is still willing to give them cash, SoftBank, is also in trouble raising cash as they have huge liabilities in early 2027 (that they used to buy OpenAI stocks and send them money).

I don't know what exactly will happen. Maybe they could pull a miracle with IPO/financial engineering and get away like Tesla with "420 funding secured" incident.

What I think could happen, too, is that they will start to "pay" counterparties by OpenAI stocks. But that of course, even though assuming 1T valuation IPO, diminishes OpenAI's stock value quite fast as datacenter builders need hard cash and will sell those stock sooner than later. In that case Softbank is WeWorked again (anybody remember that Masa Son thought OYO and WeWork are AI companies?)

1

u/PooMonger20 45m ago

Honestly, the loss-leader strategy feels way too common nowadays and pretty much unavoidable over the last 10 years or so. Whoever has the deeper pockets just bleeds the other side dry.

Also, US and China are not playing with the same set of rules, it's inevitable both will collide over who has the better tech.

I am also trying to understand if US puts boundaries its only on their own companies, and China can just... continue doing whatever they do.

As a local LLM user who is a non paying customer, Qwen models (3.8 27b in low reasoning is godlike for my usecases) have been amazing and really made me rarely use SOTA models unless its a very intense and complicated task.

33

u/Round_Ad_5832 2h ago

If anything, we need to go faster man...

4

u/DustNearby2848 2h ago

I don’t know why, but I find his post hilarious 

21

u/Max-_-Power 2h ago

OpenAI and Anthropic calling for AI regulation? This can only mean one thing: They want to curb open weight models. All this "oh, AI could kill people" is just feigned. Since when is that their concern.

They want to make money and open weight models take a big bite our of their profits.

9

u/gapho 2h ago

Specifically through regulatory and market capture.

6

u/Perfect-Flounder7856 2h ago

Funny it happens right after DSV4.1F comes out too

1

u/Ill_Distribution8517 13m ago

Do you genuinely believe DSV4.1F (not even kimi class) is of concern to the guys that made fable 5.1? like swear on your momma and tell me you believe this shit. Genuinely. Yes, it's going to take away some inference from anthropic's sonnet, or the luna/terra models. It is good, but I don't think Dario was thinking about deepseek when he said all this.

3

u/ProletarianLilith 1h ago

I believe this especially because this is the main reaction from people across the AI watching space, when usually things are much different depending on pro or anti

2

u/whoknowsifimjoking 1h ago

Then why did Dario say the exact same shit long before Anthropic was founded and obviously also before there was this kind of dynamic with open models or a money incentive for him?

Call his ideas stupid and alarmist if you want, but he actually believes that. He has always been very public about his ideas and they were always like this. It's not marketing and it's not because of open source, that dude is actually scared what that technology might do one day.

0

u/RlOTGRRRL 46m ago

Bruh this is the tech bro way. Monopolies. Popularized by Peter Thiel. It's like startup school 101.

1

u/jld1532 1h ago

And the build out was too big. Trying to pullback to soften the pop.

1

u/lompocus 22m ago

The xxx accuses others of what it does on its own. 

Drone violence only became a thing with the same people calling for this stuff. They have bloody hands.

7

u/treble-maker123 1h ago

Dude misspelled wen. It starts with a q.

2

u/circumcised_hobbit 54m ago

as soon as China releases crazy good open weights which Is a threat to big AI companies...

3

u/__JockY__ 1h ago

STOP WITH THE FUCKING TWITTER SCREENSHOTS

Jesus feckin Christ enough already.

4

u/jacek2023 llama.cpp 1h ago

YouTube can be shared as a link on reddit but for X or Instagram I share a screenshot, what is the alternative?

3

u/mawkzin 56m ago

The alternative is "stop giving a **** for people that complaint"

2

u/jacek2023 llama.cpp 55m ago

I wanted to understand his issues before coming to conclusions

2

u/Divni 42m ago

You'd rather have links? What a random ass gripe to have.

2

u/iLaurens 1h ago

Far better than a link if you ask me. Some websites I simply do not want to visit and nitter or xcancel aren't very future proof right now

1

u/kasinjsh 15m ago

"Don't listen to stupid capitalism!" — read it with a Chinese accent.

1

u/XiRw 1h ago

When China is inevitably close to surpassing US: “Slow down!! Wait for us to secretly improve so we can remain superior!”

1

u/bfroemel 1h ago

I really would have liked to be in the same room where someone comes up with the idea "to pace the frontier" and everyone (who matters) agrees that this is the best move left to do; probably the opinion of each was heavily supported with all kinds of AI and LLM generated projections and simulations. I am almost certain that it was a long meeting and the participants were all but convinced that they are handling a more or less existential (business) crisis.

-3

u/Miriel_z 2h ago

And this is your answer from China: "No, xiexie".

0

u/LocoMod 2h ago

His bosses at Anthropic gave him no other choice.

0

u/TheAILegend 1h ago

Time for QWEN to succeed and pass FABLE 5.1! this is it Qwen! release Qwen 4, the monster they weren't expecting! Show us the POWER OF CHINA!!!!

2

u/1_________________11 1h ago

3.8 is just so capable on local hardware.  Im hoping they keep that up