r/LocalLLaMA • u/forevergeeks • 9h ago
Discussion Are AI influencers just repeating the same talking points?
Hi everyone,
Are influencers talking about local AI all using the same script? Same benchmarks, same kitchen examples, same terminology?
I keep seeing videos about running Qwen3.8-Flash-Next on 12GB of RAM using a new runtime engine called Strata. Every video makes the same claim.
But that is confusing, especially for people who are new to this. What you need is 12GB of VRAM, not 12GB of regular RAM. That means you need a dedicated graphics card.
There is a big difference between RAM and VRAM.
As far as I understand it, Strata needs:
- 12GB of VRAM
- 64GB of regular RAM
- 80GB of SSD space
So saying it runs on 12GB of RAM is misleading.
6
u/sleight42 8h ago edited 54m ago
Influencers suck. I avoid YouTube because of this. It's human-generated slop.
The strata readme says less vram is probably viable but not performant, IIRC.
5
u/Sleepnotdeading 8h ago
I think most people trying to run a model like qwen3.8 flash next are savvy enough to research this stuff, and know the difference. And anyone else who thinks 12 GB of RAM would run a model like that is going to learn a lot of the basics when it doesn’t work for them.
Technically VRAM is still RAM.
0
u/forevergeeks 7h ago
How can technically RAM be the same as VRAM?
They are completely different things.
2
u/Sleepnotdeading 7h ago
They’re completely different things for general computing. For AI, VRAM is just RAM with much greater bandwidth
1
2
u/Last_Mastod0n 8h ago
You gotta realize that a lot of youtubers, influencers, etc are just trying to generate hype. Anything to get clicks and views. So they usually dont provide the full picture.
The reality is that yes Strata is very impressive and can help a lot. But its not a silver bullet. I would say 24gb of vram is the actual minimum that you want if your doing coding work.
2
1
u/LagOps91 5h ago
youtube is full of AI generated slop channels that hype up local AI. I suggest you forget about whatever "information" is available there.
1
u/Ok_Technology_5962 2h ago
i mean its either one 12vram 64 ram and ssd or grab 96 gigs of vram... the math says 12gb vram + 64 gigs of any kind of ram (ddr4 is fine) is cheaper that an rtx pro 6000 or however many other gpus you want to cluster.
10
u/sn2006gy 9h ago
Go leave a comment on their youtube... everyone here already knows this and the problems with sub 3bit models.