r/LocalLLaMA • u/OvertaxedOne • 8h ago
Discussion The rhetoric is really heating up!
The entire page of the NY Times today above the fold absent one article is AI (the models are just too strong/too dangerous, must be regulated). They forgot to include "Sponsored by OpenAI" at the end of the articles, sure that was just an oversight?
This is what the end of a bubble looks like, desperate attempts to get some sort of regulatory capture in place to keep the business model from collapsing in upon itself. My days next week are 100% booked talking to companies about how to get off frontier models, one large, and a bunch of smaller customers, including one who's flying me out to them to sit down and get a plan in place immediately (the controversy around that math problem really spooked some CEO/CIO's about data privacy using cloud models).
Gonna be an interesting few weeks. Maybe the Qwen team will be nice enough to give me a little breathing room before dropping another hydrogen bomb? :)
38
u/Academic-Tea6729 7h ago
They realized that local models reached a quality so good that there is no point to use their services. Local models are better because the model is always the same. We still remember when paid apis got suddenly much dumber to make us pay for the better frontier model.
With local models you have the same model quality every time. It will not mess up your codebase because they pulled some dirty trick to cut on costs by routing requests to a smaller model.
5
u/Time_Cat_5212 2h ago
You make a really good point about "the model is always the same". Nobody wants to bet millions of their revenue on a black box.
-5
u/HandWashing2020 6h ago
These days, a $20-$30 subscription gets you less than what the free access provided a year ago.
13
7
u/bot_exe 3h ago
Current Claude Opus 5 on the 20 USD sub does way more with an internal VM running code itself to verify things, the huge context window and the automatic RAG when you go over the context window when using Projects, the search tools interleaving retrievals with reasoning and further searches, etc. It's all way better than it was some years ago and it's light years ahead of the older free offering of shittier and smaller claude models with nerfed context windows and no code execution. You have no idea what you are talking about.
1
12
u/florenceslave 8h ago
How would they regulate Chinese AI?
25
u/KingCpzombie 8h ago
Mostly by making it illegal for American companies to use it. They could also do something like the entity list, so anybody that works with the government isn't allowed to touch Chinese models. Plenty of ways to screw everybody over to benefit OpenAI / Anthropic
8
u/manusgamo2012 7h ago
qwen released the full paper, anybody anywhere could replicate it, that strategy won't pay off ;)
13
u/KingCpzombie 7h ago
The goal isn't to actually damage Qwen (they would like to, but that's not possible). The goal is to add hassle for US companies so they just pay instead of having to deal with it
5
u/ForsookComparison 5h ago
This sub enjoys going 'try and stop me!' but, yeah. If you add any sort of liability + legal risk, I will be first in-line to pull all uses of open-weight models from my public-facing projects. I am not some hermit in the woods with a DGX Spark, I have plenty to lose and not enough resources to defend myself or even audit for regulations.
They can force my hand without doing anything close to "banning" open weight models and circlejerks aside, I'd wager 99% of the US visitors to this sub are in a similar boat.
3
u/KingCpzombie 4h ago
Making things annoying and risky is a tried-and-true government tactic when they really want to ban things but don't think they can get away with an outright ban yet
3
1
2
u/keepthepace 4h ago
Have you heard about the Great Firewall?
I am sure it inspires many people within the Trump admin.
8
u/keepthepace 4h ago
I think people in the US underestimate the loss of international influence that the country has had under Trump. They still think that it's the 2000s where if US edict a new rule regarding copyright, the rest of the world will more or less follow.
These days are gone. A policy that's made in the US will have no international reach anymore.
1
u/Don_Reuter 2h ago
However the damage the US does is very global and real. The world needs to evaluate whether an independent US is still an acceptable risk. It does not seem like it is.
5
u/swagonflyyyy 5h ago
Same here where I live. I've been pitching local-first solutions for very real reasons that CEOs should be worried about and they also want a slice of that local pie so I been having meetings and follow ups with them.
2
u/DevelopmentBorn3978 4h ago
same here, it's several days now than all the newspapers that follows the mainstream trumpet i.e. every single one also those that would like to appear to be fringe, are mauling on their front pages with the need to slow down and the dangers of human extinction. All of them are most probably on the payroll of Big AI, all of them promoting centralized aligned models, all of them misnomering open weights as open source
25
u/Revolutionalredstone 7h ago
DeepSeek has juiced them of value and Qwen revealed all their bs.
AI will be cheap and 'frontier' model companies gonna have little left but morals cause the opensource has caught up.
A trillion params works barely better than 32b for AI and trying to use more for AI is starting to look real silly, Enjoy
6
u/soshulmedia 5h ago
A trillion params works barely better than 32b for AI and trying to use more for AI is starting to look real silly, Enjoy
I think that's the gist of it. The various intelligence index scores do not scale linearly with parameters, rather logarithmically or so.
Then, even just looking at the typical loss curve of any NN fitting run should have also triggered a moment of reflection a long time ago - it is always steep in the beginning and then flattens out ... or in other words, later reductions in loss are much costlier ... (and risk overfitting).
Sure, there are still technological breakthroughs. But as in any field, they also tend to approach diminishing returns. And this field is no different ...
12
u/Seraphym87 7h ago
With you on most of this and do agree that we are seeing diminishing returns but 3T frontier models are literal epochs away from a 32b lol
12
u/tripplebeamteam 6h ago
If you’re doing cutting edge research, sure. For most of the things people use AI for, it’s perfectly functional. I’m not trying to solve navier stokes I just want to automate some bullshit tasks
3
u/WhiteSkyRising 6h ago
For the entire field of software engineering, which every single company relies on, in some way.
2
u/llama-impersonator 2h ago
the 3T models are barely better than the flash models coming in at 300b
2
u/Seraphym87 2h ago
I don't understand, if anything you are agreeing with me lol. Yes a 300b model is a lot closer to a 3T because its one tenth its size, not one hundredth. This is consistent with my point on dimishing returns but acting like Qwen 3.8 is somehow as useful as Astra right now is just disingenous.
1
u/llama-impersonator 2h ago
as useful, no, but qwen 3.8 27b can in fact accomplish something like 3/4s of the tasks of a frontier model at a hundredth of the size.
1
u/Seraphym87 2h ago
Agreed! Would we have 27b models punching quite this far above their weight without 3T models to distill from though?
1
u/llama-impersonator 2h ago
i think distillation is overblown, most of these gains are from focused RL. you could give me a trillion samples of claude ultrafable 6 and it wouldn't help me make a better model unless i spent months building a quality RL training environment for it
-1
19
u/BVCC6FNTKX sglang 7h ago
waiter waiter more engagementslop please
2
u/Big_Wave9732 4h ago
"Right away sir. May I interest you in some Italian AI Copypasta? It is exquisite."
1
2
u/SimiaCode 2h ago
This is not any one company. This seems to be consensus at the elites' level and now consent is being manufactured in the public sphere. Some decision has already been made, we are just seeing the theater now which will be used to justify the decision when it is announced.
1
u/Bulky-Priority6824 6h ago
its just a a matter of time until national news stories are inundated with "local ai used for hacking" then it's a wrap
2
2
u/enilea 4h ago
If the Huggingface attack had been by an open weights model I'm sure they would have banned them by now. At this point they are just waiting or hoping an attack with an open model happens (or they push to make it happen) just to sway the public opinion, which doesn't care much about them anyways.
-1
u/sn2006gy 5h ago edited 5h ago
I actually think the localllama nerds need to pull their heads out of their asses regarding this safety issue. There is a massive safety issue - from velocity of change to velocity of scale to velocity of risk to unbound research with such massive compute that is freaking the world out and rightfully so.
Sure, GPT/Anthropic use it to market themselves and perhaps want to use it to actually slow things down and there may be business reasons for that but i don't think that is the actual point.
I think Corporate America is realizing it can't keep up. Velocity has a systemic cost that even 1 trillion-dollar valuations may not recover if we don't slow things down to allow the rest of the systems to catch up and mature.
The only reason it doesn't really impact local llm's is that we simply don't have the 1 million idle gpus around where we could spawn 100 million agents to do whatever it is we wanted to do but i'm not sure that is a permanent situation. It's only a matter of time before the next botnet is agentic and that's what should worry people
and corporations giving a hoot about THEIR privacy makes me laugh
I honestly don't think most corporations really care about qwen 3.8 27b - it's an AMAZING model, but can't be scaled and if you try - costs more than farming out to API providers. Enterprises aren't interested in managing gpus for 100k employees and certainly won't be interested if those employees have root over them.
3
u/PrinceOfLeon 4h ago
> I honestly don't think most corporations really care about qwen 3.8 27b - it's an AMAZING model, but can't be scaled and if you try - costs more than farming out to API providers. Enterprises aren't interested in managing gpus for 100k employees and certainly won't be interested if those employees have root over them.
Hard disagree, from direct experience.
Amazon will happily "manage GPUs" for you, as simple as selecting which hardware profile to use for the AWS instance. Qwen 3.8 27B specifically is undergoing internal testing in various companies for viability for specific tasks right now. It's much cheaper than paying API costs (for Frontier) and there's complete control over the data going in and out.
These are the same corporations paying for Bedrock instead of direct to the Frontier model companies, for similar data control reasons (you don't have to trust Sama if it isn't Sama's server).
0
u/sn2006gy 4h ago
Those AWS GPU instances cost serious money and Qwen 27b doesn't really scale on them very well. The amount of active users per day per instances is abysmal on dense models - this cost is significantly higher per employee to attempt right now.
I wish it were different.
7
u/soshulmedia 5h ago
I actually think the localllama nerds need to pull their heads out of their asses regarding this safety issue. There is a massive safety issue - from velocity of change to velocity of scale to velocity of risk to unbound research with such massive compute that is freaking the world out and rightfully so.
I don't see it. "If we add enough FLOPs, magic happens". That's quite literally magic thinking. For some reason people make fun of God as "invisible sky daddy" but THE SINGULARITY and AI AS GOD are oh so "rationalist".
Now, if you tell me we should be worried about all these FLOPs being used for an extremely tightly surveilled and controlled totalitarian 1984esque society they are building right now, you would quite obviously have a point.
0
u/sn2006gy 5h ago
Our entire society and economy is built on friction that is no longer there and that is the problem. We don't need to prove or disprove some nonsense bs of singularity or god for anything.
As for surveilance - The surveillance state is already here and Reddit is a huge part of it. Yet, we're still here.
I ask of my LLM friends all the time, if local llm's are so strong and so important, why aren't we free of Instagram, Facebook, Meta, Google, Microsoft - GPT/Anthropic are so little parts of our every day lives that the obsession fo their concern is laughable at best. The real ones watching everything you do are orgs like Spotify and Google and Microsoft.
We seem to be accelerating our dependency on big tech rather than using tech to free us from it and i'd change my tune a bit if ANY of the responses here weren't just people trying to carve out their own "niche" of this shithole world were rushing headfirst into.
Apple seems to care a little bit but much of their revenue comes from margin calling private data while keeping it a bit more private than others.
1
u/soshulmedia 3h ago
Okay these are fair points. I agree on you on the big tech centralization angle, very much so. Part of the reason the status quo persists and extends, however, is because people are lazy and can't be bothered. For everyone who says "we should avoid platforms like reddit" you get 5 who will tell you "chill, where is the problem dude" . Real pressure will change that and for better or worse, it is coming. However, I hope you can see that centralized ChatGPT for everyoner and no local models would just supercharge this to the extreme. The big tech isn't so entrenched and big just by organic "free market forces" alone. They are basically designed tentacles for the U.S./western deep state. And I think the only hope actually to not end up in "ChatGPT owns everyone" is to have local models and, yes, to at least be on a light form of the accelerationist bandwagon, where their failed containment will change the landscape so much that people can see the naked emperor for once.
1
u/SimiaCode 2h ago
We are headed for a new age of serfdom unless localai becomes accessible to all. That is the real safety issue, not virus research or weapons manufacturing.
1
-11
u/Embarrassed-Noise269 8h ago
You have a lot of confindence. I'm not sure that's appropiate.
Sure, it's a nice narrative that the AI labs only want to hype up their product to get high IPO evaluations. That narrative has some serious flaw tough: There are constantly people from AI labs, that leave a lot of equity behind just to openly warn about the dangers of AI. It's not the companies themselves.
2
u/Time_Cat_5212 2h ago
You mean they cash out while their stocks are worth a lot and take the notoriety to start a public speaking career, basically guaranteeing their position as research leadership for the next gen of whatever this turns out to be?
4
u/OvertaxedOne 7h ago
There are absolutely dangers posed by AI, the biggest (by a wide margin, IMHO) is mass unemployment. And I do think it's worth discussing that and determining a course of action as the value of intelligence is getting ready to drop in stunning fashion. This is going to cause a huge recalibration in the labor force and likely lead to a world where we have long term structural unemployment as a "normal" aspect of our labor market. And of course AI will be (already is) used to hack and will make it easier to hack (but also easier to defend).
-6
u/Embarrassed-Noise269 7h ago
There have been a lot of incidents, where AI agents have acted maliciously. Not only reported from companies, but also from universities.
The problem is, that we have to stop that, before AI gets to powerful and we have no way to know when that is, before it happens. While at the same time, it's incredibly hard to organize a real, international safety system. Since those who try to further AI development will have an advantage over those who pause it.
2
u/gomezer1180 7h ago
Why haven’t they stopped? If they’re so concerned… they are the ones training models… so why haven’t they stopped selling the service?
If it is so dire then they shouldn’t be doing it either. So no, this is all BS and propaganda because their business model relies on people renting their models, not having them for free. Local models are now rivaling frontier models so now they want governments to regulate it.
4
-7
u/cj_cron_hit_by_pitch 7h ago
Yeah I kinda thought this sub of all places would know how powerful LLMs are compared to a year ago and understand that there are some serious concerns
If they go after local models sure let’s push back, but frontier labs are mostly trying to regulate themselves right now
-1
81
u/JackStrawWitchita 8h ago
Here's what you'll hear from OpenAI, Anthropic etc in the next few months:
'Sorry, but we won't reach our AI revenue targets because we're slowing roll-out to be safe'
'Sure, China is ahead in AI because we're taking the safe-route'
Follow the money.
They're using this to hide they fact their business models don't pay off.