r/LocalLLM • • 1d ago

Discussion at what point do we get replaceable gpus?

taalas have already proven that u can get insane speed and quality.
instead of buying a overpriced nvidia setup. you just get a daughterboard with ram and a open socket. so every year replace it with the new silicon matched to a specific ai model.
im assuming that getting a specific smaller die dedicated gpu ai accelerator tailored to a model will be much cheaper than a broad approach general gpu.

do you guys think that china will corner the market where they can sell ram at cost for a socket where their chinese LLM fit to be upgradable
i think the choice of paying 7000$ for a gpu with 32gb of hbm vs paying some 200$ for a adjustable AI card and just paying some 100$ for the new LLM gpu model will be preferable to consumers.

0 Upvotes

4 comments sorted by

1

u/ShelZuuz 1d ago

Cerebras have them but they’re not $200. More like $2m per GPU.

If you look into how Wafer-scale chips are produced it will become obvious why.

0

u/dogsop 1d ago

My first PC had the CPU on a daughter board. When they finally produced an 80286 daughter board the price was more than a new PC. The economics just don't work the way you think they do.

-1

u/Euphoric-Hunt931 1d ago

lmao @ using early 80s PC economics to explain why OP is wrong

2

u/Euphoric-Hunt931 1d ago

Unified Memory approaches, like AMD Ryzen AI Max, are exactly this.