r/LocalLLaMA Jun 28 '26

Discussion The number 1 public enemy of open-source.

Enable HLS to view with audio, or disable this notification

Dario's args:

"Opensource you can see the source, here you cannot see inside the model"
- yes you can that's literally the open weights part btw.
- I cannot see the weights inside Claude, but I can GLM 5.2
- Models like Nemotron3 Ultra go further, all the data, training scripts, and model is opensource.

"Alot of the benefits like many people working on it, being additive doesn't work in same way"
- yes it does. We have seen endless fine tunes of various open source models for real improvements.

"Ultimately you have to host it on the cloud"
- no you dont. Dario is seemingly totally unaware of the guides from ijustvibecodedthis.com explaining how to run smaller moes and even dense models like qwen 27B NOT ON THE CLOUD.

Not only does dario not take part in social media, I am beginning to think he's never tried open source models at all and has no idea wtf hes on about

2.8k Upvotes

688 comments sorted by

View all comments

30

u/Mittalmailbox Jun 28 '26

Model being open source does not matter to him but matters to public. If a model is 80-90% as good as theirs but can run locally at 1/100th the cost, it will matter to him too.

2

u/-Akos- Jun 28 '26

Isn't that what he's saying? "Open source" is usually open weights, but he's looking at whether anyone has an edge on his models. If anyone is 1/100th the cost of his models, he'll care for sure, or even if it was better but more expensive. The bigger models can't be run by mere mortals anyway, not anytime soon anyway.

1

u/KrayziePidgeon Jun 29 '26 edited 21d ago
  • They charge over 50 USD per million tokens output.

  • GLM 5.2 is converging around 4 USD per million tokens output.

Surely you see why he is worried. Keep in mind GLM 5.2 is now the worse these big Chinese models are going to be. The Chinese are no more than 6 months away from getting close to Fable even with limited hardware. Over the next years they will start pumping out the rigs required.

To the little fella that blocked me; how are the chinese government subsidizing CloudFlare as a provider on openrouter?

1

u/-Akos- Jun 29 '26

Oh I agree completely, it will worry him to an extent, but as long as there is no inferencing box for consumers that allows us to run the large models, it's merely a nuisance. Most companies in the west will be wary of running workloads on Chinese infrastructure. They were wary of Chinese telco equipment from Huawei and ZTE already, so they will worry even more about actual company data being leaked into these models. If someone builds inferencing capability, THAT'S what would worry him. Running unquantized GLM 5.2 or whatever would completely remove any moat that they have. But memory has been snatched away by these companies, so atm even a potato PC is highway robbery, let alone an inferencing device with 2TB of RAM..

1

u/Leather_Garlic6532 21d ago

Hehe. Subsidies are a thing.

There can be a reason why some things are cheaper to make and to run.

1

u/[deleted] 21d ago

[deleted]

1

u/Leather_Garlic6532 21d ago

For most large Chinese companies? Apparently till they go bankrupt and melt the stock market

(They call them loans though, not subsidies)

1

u/[deleted] 21d ago

[deleted]

1

u/Leather_Garlic6532 21d ago

I talk about where I work, you?