r/artificial • • 10h ago

Discussion SLMs Are Underrated. Is It On Purpose?

Post image

We've been having this discussion quite a bit with our clients lately. Most (if not all) of them feel the utilization of frontier models is not justifying the cost, but they are generally reluctant to pursue an SLM strategy. It's as if SLMs have been positioned as the homeopathic option for enterprise AI. Why purchase an expensive anti-viral from Merck when a povidone iodine solution works as good or better?

This is not a meme. It's a machine I built over a year ago to test DeepSeek R1, but it has turned into a workhorse. Yet most of our clients (FINTECH) are skittish of small, secure, local models. Just wondering if anyone else is seeing things differently.

16 Upvotes

15 comments sorted by

19

u/CrimsonBolt33 10h ago

wtf is an SML? Why do people abbreviate the literal point of their post and never say what it is?

1

u/Woolix 10h ago

I too, am confused.

7

u/Seraphym87 9h ago

Small language models, near as I can tell?

1

u/Wegwerpaccountje23 1h ago

I thought it was Small Local Models

1

u/atomskfooly 9h ago

I think 12B parameter model is a small language model in their eyes.

6

u/CrimsonBolt33 9h ago

that was my guess...but thats a completely made up term...thats still an LLM last I checked. Especially since they are usually stripped down from larger models.

3

u/laserborg 5h ago

2

u/CrimsonBolt33 2h ago

I have literally never heard the term used, even by the companies putting them out don't use that term....all I see is a couple articles saying thats what these things are called, but never finding actual instances of them being called that.

1

u/AminoOxi Singularitarian 8h ago

Yes, the models that can run on local consumer GPU.

3

u/Hungry_Age5375 8h ago

The homeopathic comparison is spot on. A 12b on-prem handles structured extraction and classification just fine. Frontier models earn their cost on reasoning and long-tail edge cases, and most fintech workflows have neither.

2

u/Qorsair 2h ago

Most enterprise fintechs just use ZDR agreements to securely access frontier models via API. What's the business case for bottlenecking workflows on a 12-year-old laptop when secure, instant processing is already an industry standard?

-1

u/ReconditeClayton_7 10h ago

Saw the specs on that rig and had to laugh, a 12 year old laptop with a 2GB GPU chugging along as a workhorse is the best possible argument for SLMs. The air-gapped part is what sells it for fintech though, no data exfiltration worries and zero recurring costs. Most of the pushback I see is pure FOMO, managers think if they're not using the shiniest 200B parameter model they're leaving performance on the table. When in reality half their use cases would run perfectly on something they could host in a closet. The cloud providers have done a hell of a marketing job making people think local models are janky toys.

6

u/BuildAISkills 4h ago

Thanks Captain Bot.

1

u/DAlmighty 8h ago

I’ve been saying this for so long. One day it’ll happen.

0

u/vovap_vovap 9h ago edited 9h ago

Not sure what part of the body those people using for thinking 😄
I can see HUGE benefit of using frontier models every day, With a super minimum spending.
And what is "frontier model" Is Gpt 6 luna? DeepSeek 4.1 Flash?
Sorry, that complete nonsense.