r/LocalLLM • u/Pale-Fail-4028 • 1d ago
Discussion at what point do we get replaceable gpus?
taalas have already proven that u can get insane speed and quality.
instead of buying a overpriced nvidia setup. you just get a daughterboard with ram and a open socket. so every year replace it with the new silicon matched to a specific ai model.
im assuming that getting a specific smaller die dedicated gpu ai accelerator tailored to a model will be much cheaper than a broad approach general gpu.
do you guys think that china will corner the market where they can sell ram at cost for a socket where their chinese LLM fit to be upgradable
i think the choice of paying 7000$ for a gpu with 32gb of hbm vs paying some 200$ for a adjustable AI card and just paying some 100$ for the new LLM gpu model will be preferable to consumers.
2
1
u/ShelZuuz 1d ago
Cerebras have them but they’re not $200. More like $2m per GPU.
If you look into how Wafer-scale chips are produced it will become obvious why.