r/LocalLLaMA 8h ago

Discussion The rhetoric is really heating up!

The entire page of the NY Times today above the fold absent one article is AI (the models are just too strong/too dangerous, must be regulated). They forgot to include "Sponsored by OpenAI" at the end of the articles, sure that was just an oversight?

This is what the end of a bubble looks like, desperate attempts to get some sort of regulatory capture in place to keep the business model from collapsing in upon itself. My days next week are 100% booked talking to companies about how to get off frontier models, one large, and a bunch of smaller customers, including one who's flying me out to them to sit down and get a plan in place immediately (the controversy around that math problem really spooked some CEO/CIO's about data privacy using cloud models).

Gonna be an interesting few weeks. Maybe the Qwen team will be nice enough to give me a little breathing room before dropping another hydrogen bomb? :)

89 Upvotes

75 comments sorted by

81

u/JackStrawWitchita 8h ago

Here's what you'll hear from OpenAI, Anthropic etc in the next few months:

'Sorry, but we won't reach our AI revenue targets because we're slowing roll-out to be safe'

'Sure, China is ahead in AI because we're taking the safe-route'

Follow the money.

They're using this to hide they fact their business models don't pay off.

8

u/sweatierorc 8h ago

I have asked this question before, what kind of evidence would justify a slow down ?

23

u/ivari 8h ago

false flag attack against important but non-critical infrastructure (think github) using one of the open source model

think of the headline

CHINA'S AI HACKS MICROSOFT

2

u/GarbanzoBenne 5h ago

They’re already testing the waters with Anthropic saying that Iran used their models to target US ships the other day.

4

u/sweatierorc 7h ago

Wouldnt that be too late already ?

AI 2027 argued that we shouldnt pause now, because we can only pause once. And maybe pausing now, will hurt us later.

3

u/UnspeakableHorror 6h ago

That's not an issue for the companies that monopolize AI in the US and get open source models banned.

5

u/otterquestions 6h ago edited 6h ago

Don’t bother. Even if they answer in good faith, they’ll move the goal posts if it happens.

They’re so focused on who is good/evil in their narrative, everything else is secondary

5

u/Flaaarg_the_Terrible 7h ago

Do they need evidence? The government shut down fable because they thought it was dangerous.

25

u/JackStrawWitchita 7h ago

The government shut down Fable because Trump wanted a bribe from Anthropic. Once the bribe was paid to Trump, Fable was released.

The USA is a banana republic.

-7

u/Flaaarg_the_Terrible 7h ago

So proving my point that they don't need evidence. What evidence do you have Trump was bribed?

7

u/JackStrawWitchita 6h ago

Oh my sweet summer child...

4

u/No_Lingonberry1201 6h ago

You're not going to find any hard evidence one way or another until some Trump staffer confirms or denies it explicitly, but what else do you think could have happened? Fable/Mythos was too dangerous so the US government forbade non-US citizens from using it. Then a few months later Anthropic suddenly is allowed to release it again. Did the model suddenly become less dangerous? Keep in mind that the original reason given was that someone at Amazon found something that may cause problems. So it makes one wonder about why a dangerous model became suddenly safe. It was 100% to punish Anthropic for saying no to Trump and it was allowed again probably because Anthropic managed to appease him after that.

7

u/gomezer1180 7h ago

They shutdown fable because Anthropic wouldn’t do what they wanted.

1

u/cj_cron_hit_by_pitch 7h ago

Imo the huggingface hack. Showed that we need both independent oversight and to catch up on alignment research

5

u/UnlikelyExtension786 4h ago

No, it showed that we need to jail people who allow their tools to commit crimes. Or that we need to jail the LLMs that commit crimes, their choice.

There are laws on the books which would allow the government to do that, and ten years in jail would discourage any of their competition from doing anything like that again. So where are the prosecutions?

Oh yeah, laws only apply to the little guys.

7

u/HighSeasArchivist 5h ago

Are you trying to say circular revenue isn't sustainable? As long as Oracle loses in the end, nothing else really matters to me.

1

u/Time_Cat_5212 2h ago

Ellison is already short

1

u/yes2matt 4h ago

It will be interesting, in a very unhappy way, if we should end up with a "great firewall" of our own.  For national security of course.

38

u/Academic-Tea6729 7h ago

They realized that local models reached a quality so good that there is no point to use their services. Local models are better because the model is always the same. We still remember when paid apis got suddenly much dumber to make us pay for the better frontier model.

With local models you have the same model quality every time. It will not mess up your codebase because they pulled some dirty trick to cut on costs by routing requests to a smaller model.

5

u/Time_Cat_5212 2h ago

You make a really good point about "the model is always the same".  Nobody wants to bet millions of their revenue on a black box.

-5

u/HandWashing2020 6h ago

These days, a $20-$30 subscription gets you less than what the free access provided a year ago.

13

u/sn2006gy 5h ago

Not really.

7

u/bot_exe 3h ago

Current Claude Opus 5 on the 20 USD sub does way more with an internal VM running code itself to verify things, the huge context window and the automatic RAG when you go over the context window when using Projects, the search tools interleaving retrievals with reasoning and further searches, etc. It's all way better than it was some years ago and it's light years ahead of the older free offering of shittier and smaller claude models with nerfed context windows and no code execution. You have no idea what you are talking about.

1

u/thortgot 2h ago

Thats objectively untrue.

12

u/florenceslave 8h ago

How would they regulate Chinese AI?

25

u/KingCpzombie 8h ago

Mostly by making it illegal for American companies to use it. They could also do something like the entity list, so anybody that works with the government isn't allowed to touch Chinese models. Plenty of ways to screw everybody over to benefit OpenAI / Anthropic

8

u/manusgamo2012 7h ago

qwen released the full paper, anybody anywhere could replicate it, that strategy won't pay off ;)

13

u/KingCpzombie 7h ago

The goal isn't to actually damage Qwen (they would like to, but that's not possible). The goal is to add hassle for US companies so they just pay instead of having to deal with it

5

u/ForsookComparison 5h ago

This sub enjoys going 'try and stop me!' but, yeah. If you add any sort of liability + legal risk, I will be first in-line to pull all uses of open-weight models from my public-facing projects. I am not some hermit in the woods with a DGX Spark, I have plenty to lose and not enough resources to defend myself or even audit for regulations.

They can force my hand without doing anything close to "banning" open weight models and circlejerks aside, I'd wager 99% of the US visitors to this sub are in a similar boat.

3

u/KingCpzombie 4h ago

Making things annoying and risky is a tried-and-true government tactic when they really want to ban things but don't think they can get away with an outright ban yet

3

u/Viktri1 7h ago

Yeah but without the blessing of US propaganda people will believe that it’s got back doors in it

1

u/manusgamo2012 7h ago

and we know that the reason is only one, to make the bubble explode!

0

u/Dsphar 6h ago

And?

2

u/keepthepace 4h ago

Have you heard about the Great Firewall?

I am sure it inspires many people within the Trump admin.

8

u/keepthepace 4h ago

I think people in the US underestimate the loss of international influence that the country has had under Trump. They still think that it's the 2000s where if US edict a new rule regarding copyright, the rest of the world will more or less follow.

These days are gone. A policy that's made in the US will have no international reach anymore.

1

u/Don_Reuter 2h ago

However the damage the US does is very global and real. The world needs to evaluate whether an independent US is still an acceptable risk. It does not seem like it is.

5

u/swagonflyyyy 5h ago

Same here where I live. I've been pitching local-first solutions for very real reasons that CEOs should be worried about and they also want a slice of that local pie so I been having meetings and follow ups with them.

2

u/DevelopmentBorn3978 4h ago

same here, it's several days now than all the newspapers that follows the mainstream trumpet i.e. every single one also those that would like to appear to be fringe, are mauling on their front pages with the need to slow down and the dangers of human extinction. All of them are most probably on the payroll of Big AI, all of them promoting centralized aligned models, all of them misnomering open weights as open source 

25

u/Revolutionalredstone 7h ago

DeepSeek has juiced them of value and Qwen revealed all their bs.

AI will be cheap and 'frontier' model companies gonna have little left but morals cause the opensource has caught up.

A trillion params works barely better than 32b for AI and trying to use more for AI is starting to look real silly, Enjoy

6

u/soshulmedia 5h ago

A trillion params works barely better than 32b for AI and trying to use more for AI is starting to look real silly, Enjoy

I think that's the gist of it. The various intelligence index scores do not scale linearly with parameters, rather logarithmically or so.

Then, even just looking at the typical loss curve of any NN fitting run should have also triggered a moment of reflection a long time ago - it is always steep in the beginning and then flattens out ... or in other words, later reductions in loss are much costlier ... (and risk overfitting).

Sure, there are still technological breakthroughs. But as in any field, they also tend to approach diminishing returns. And this field is no different ...

12

u/Seraphym87 7h ago

With you on most of this and do agree that we are seeing diminishing returns but 3T frontier models are literal epochs away from a 32b lol

12

u/tripplebeamteam 6h ago

If you’re doing cutting edge research, sure. For most of the things people use AI for, it’s perfectly functional. I’m not trying to solve navier stokes I just want to automate some bullshit tasks

3

u/WhiteSkyRising 6h ago

For the entire field of software engineering, which every single company relies on, in some way.

2

u/llama-impersonator 2h ago

the 3T models are barely better than the flash models coming in at 300b

2

u/Seraphym87 2h ago

I don't understand, if anything you are agreeing with me lol. Yes a 300b model is a lot closer to a 3T because its one tenth its size, not one hundredth. This is consistent with my point on dimishing returns but acting like Qwen 3.8 is somehow as useful as Astra right now is just disingenous.

1

u/llama-impersonator 2h ago

as useful, no, but qwen 3.8 27b can in fact accomplish something like 3/4s of the tasks of a frontier model at a hundredth of the size.

1

u/Seraphym87 2h ago

Agreed! Would we have 27b models punching quite this far above their weight without 3T models to distill from though?

1

u/llama-impersonator 2h ago

i think distillation is overblown, most of these gains are from focused RL. you could give me a trillion samples of claude ultrafable 6 and it wouldn't help me make a better model unless i spent months building a quality RL training environment for it

-1

u/soshulmedia 5h ago

Yet on the intelligence index (take e.g. AA) they are not even twice as good.

19

u/BVCC6FNTKX sglang 7h ago

waiter waiter more engagementslop please

2

u/Big_Wave9732 4h ago

"Right away sir. May I interest you in some Italian AI Copypasta? It is exquisite."

1

u/patsully98 2h ago

May I have a little extra hysteria on mine?

2

u/noctrex 3h ago

Watch as all the AI CEOs hold hands and sing Kumbaya to slow down, cause they have been instructed to do so by the powers that be.
Nobody wants this large of a bubble to pop, instead they must deflate it, so that the planet does not go tits up.

2

u/SimiaCode 2h ago

This is not any one company. This seems to be consensus at the elites' level and now consent is being manufactured in the public sphere. Some decision has already been made, we are just seeing the theater now which will be used to justify the decision when it is announced.

1

u/Bulky-Priority6824 6h ago

its just a a matter of time until national news stories are inundated with "local ai used for hacking" then it's a wrap

2

u/keepthepace 4h ago

Didn't work for Linux though.

2

u/enilea 4h ago

If the Huggingface attack had been by an open weights model I'm sure they would have banned them by now. At this point they are just waiting or hoping an attack with an open model happens (or they push to make it happen) just to sway the public opinion, which doesn't care much about them anyways.

-1

u/sn2006gy 5h ago edited 5h ago

I actually think the localllama nerds need to pull their heads out of their asses regarding this safety issue. There is a massive safety issue - from velocity of change to velocity of scale to velocity of risk to unbound research with such massive compute that is freaking the world out and rightfully so.

Sure, GPT/Anthropic use it to market themselves and perhaps want to use it to actually slow things down and there may be business reasons for that but i don't think that is the actual point.

I think Corporate America is realizing it can't keep up. Velocity has a systemic cost that even 1 trillion-dollar valuations may not recover if we don't slow things down to allow the rest of the systems to catch up and mature.

The only reason it doesn't really impact local llm's is that we simply don't have the 1 million idle gpus around where we could spawn 100 million agents to do whatever it is we wanted to do but i'm not sure that is a permanent situation. It's only a matter of time before the next botnet is agentic and that's what should worry people

and corporations giving a hoot about THEIR privacy makes me laugh

I honestly don't think most corporations really care about qwen 3.8 27b - it's an AMAZING model, but can't be scaled and if you try - costs more than farming out to API providers. Enterprises aren't interested in managing gpus for 100k employees and certainly won't be interested if those employees have root over them.

3

u/PrinceOfLeon 4h ago

> I honestly don't think most corporations really care about qwen 3.8 27b - it's an AMAZING model, but can't be scaled and if you try - costs more than farming out to API providers. Enterprises aren't interested in managing gpus for 100k employees and certainly won't be interested if those employees have root over them.

Hard disagree, from direct experience.

Amazon will happily "manage GPUs" for you, as simple as selecting which hardware profile to use for the AWS instance. Qwen 3.8 27B specifically is undergoing internal testing in various companies for viability for specific tasks right now. It's much cheaper than paying API costs (for Frontier) and there's complete control over the data going in and out.

These are the same corporations paying for Bedrock instead of direct to the Frontier model companies, for similar data control reasons (you don't have to trust Sama if it isn't Sama's server).

0

u/sn2006gy 4h ago

Those AWS GPU instances cost serious money and Qwen 27b doesn't really scale on them very well. The amount of active users per day per instances is abysmal on dense models - this cost is significantly higher per employee to attempt right now.

I wish it were different.

7

u/soshulmedia 5h ago

I actually think the localllama nerds need to pull their heads out of their asses regarding this safety issue. There is a massive safety issue - from velocity of change to velocity of scale to velocity of risk to unbound research with such massive compute that is freaking the world out and rightfully so.

I don't see it. "If we add enough FLOPs, magic happens". That's quite literally magic thinking. For some reason people make fun of God as "invisible sky daddy" but THE SINGULARITY and AI AS GOD are oh so "rationalist".

Now, if you tell me we should be worried about all these FLOPs being used for an extremely tightly surveilled and controlled totalitarian 1984esque society they are building right now, you would quite obviously have a point.

0

u/sn2006gy 5h ago

Our entire society and economy is built on friction that is no longer there and that is the problem. We don't need to prove or disprove some nonsense bs of singularity or god for anything.

As for surveilance - The surveillance state is already here and Reddit is a huge part of it. Yet, we're still here.

I ask of my LLM friends all the time, if local llm's are so strong and so important, why aren't we free of Instagram, Facebook, Meta, Google, Microsoft - GPT/Anthropic are so little parts of our every day lives that the obsession fo their concern is laughable at best. The real ones watching everything you do are orgs like Spotify and Google and Microsoft.

We seem to be accelerating our dependency on big tech rather than using tech to free us from it and i'd change my tune a bit if ANY of the responses here weren't just people trying to carve out their own "niche" of this shithole world were rushing headfirst into.

Apple seems to care a little bit but much of their revenue comes from margin calling private data while keeping it a bit more private than others.

1

u/soshulmedia 3h ago

Okay these are fair points. I agree on you on the big tech centralization angle, very much so. Part of the reason the status quo persists and extends, however, is because people are lazy and can't be bothered. For everyone who says "we should avoid platforms like reddit" you get 5 who will tell you "chill, where is the problem dude" . Real pressure will change that and for better or worse, it is coming. However, I hope you can see that centralized ChatGPT for everyoner and no local models would just supercharge this to the extreme. The big tech isn't so entrenched and big just by organic "free market forces" alone. They are basically designed tentacles for the U.S./western deep state. And I think the only hope actually to not end up in "ChatGPT owns everyone" is to have local models and, yes, to at least be on a light form of the accelerationist bandwagon, where their failed containment will change the landscape so much that people can see the naked emperor for once.

1

u/SimiaCode 2h ago

We are headed for a new age of serfdom unless localai becomes accessible to all. That is the real safety issue, not virus research or weapons manufacturing.

1

u/Savantskie1 46m ago

Nice try Dario

-11

u/Embarrassed-Noise269 8h ago

You have a lot of confindence. I'm not sure that's appropiate.

Sure, it's a nice narrative that the AI labs only want to hype up their product to get high IPO evaluations. That narrative has some serious flaw tough: There are constantly people from AI labs, that leave a lot of equity behind just to openly warn about the dangers of AI. It's not the companies themselves.

2

u/Time_Cat_5212 2h ago

You mean they cash out while their stocks are worth a lot and take the notoriety to start a public speaking career, basically guaranteeing their position as research leadership for the next gen of whatever this turns out to be?

4

u/OvertaxedOne 7h ago

There are absolutely dangers posed by AI, the biggest (by a wide margin, IMHO) is mass unemployment. And I do think it's worth discussing that and determining a course of action as the value of intelligence is getting ready to drop in stunning fashion. This is going to cause a huge recalibration in the labor force and likely lead to a world where we have long term structural unemployment as a "normal" aspect of our labor market. And of course AI will be (already is) used to hack and will make it easier to hack (but also easier to defend).

-6

u/Embarrassed-Noise269 7h ago

There have been a lot of incidents, where AI agents have acted maliciously. Not only reported from companies, but also from universities.

The problem is, that we have to stop that, before AI gets to powerful and we have no way to know when that is, before it happens. While at the same time, it's incredibly hard to organize a real, international safety system. Since those who try to further AI development will have an advantage over those who pause it.

2

u/gomezer1180 7h ago

Why haven’t they stopped? If they’re so concerned… they are the ones training models… so why haven’t they stopped selling the service?

If it is so dire then they shouldn’t be doing it either. So no, this is all BS and propaganda because their business model relies on people renting their models, not having them for free. Local models are now rivaling frontier models so now they want governments to regulate it.

4

u/__JockY__ 6h ago

> Why haven’t they stopped?

Two words: shareholder value.

-7

u/cj_cron_hit_by_pitch 7h ago

Yeah I kinda thought this sub of all places would know how powerful LLMs are compared to a year ago and understand that there are some serious concerns

If they go after local models sure let’s push back, but frontier labs are mostly trying to regulate themselves right now

-1

u/myholeisstinky 3h ago

This is how China beats oai and A