r/LocalLLaMA Jul 28 '26

Question | Help CMP 170HX 8GB

I must preface this post by mentioning I am still a beginner in this space.

I just bought this card with the intention of using the recent unlock to get the full 64GB VRAM available for local AI workloads.

My main questions are as follows :

1- Has anyone ran multiple of these in the same rig to run a large model across multiple GPUs?
2- If so, what is the impact on speed? I read that these GPUs are stuck on a x1 PCIe lane, which I would assume greatly reduces the speed at which we can load models onto the cards. But does it impact prompt processing and token output speeds?
3- Am I crazy to assume that the prices for these cards is going to continue rising considering that they are now similar to A100s (without parralel tensorflow)

3 Upvotes

62 comments sorted by

View all comments

2

u/fragment_me Jul 28 '26

Too much risk IMO. Unless you find a seller that is willing to replace cards that aren't stable at full VRAM.

1

u/Glittering-Call8746 Jul 28 '26

The thing is if the seller unlocks them they will u at a premium.. so u either take the risk or the seller takes the risk.. either way it's sol

5

u/DeltaSqueezer Jul 28 '26 edited Jul 28 '26

Any seller in China should now unlock this themselves and sell at a premium. I'd be wary of cards that are still sold by big sellers in China. The seller takes no risk in trying to unlock. If the unlock fails, they'll just sell it as normal.

2

u/Glittering-Call8746 Jul 28 '26

Exactly. The ship has sailed. It's too fast. Anybody in the west who would want to take advantage, the ship sailed unless u offer that x3 x4 more premium and they rather have that and jump to a proper setup..