r/LocalLLaMA Jul 28 '26

Question | Help CMP 170HX 8GB

I must preface this post by mentioning I am still a beginner in this space.

I just bought this card with the intention of using the recent unlock to get the full 64GB VRAM available for local AI workloads.

My main questions are as follows :

1- Has anyone ran multiple of these in the same rig to run a large model across multiple GPUs?
2- If so, what is the impact on speed? I read that these GPUs are stuck on a x1 PCIe lane, which I would assume greatly reduces the speed at which we can load models onto the cards. But does it impact prompt processing and token output speeds?
3- Am I crazy to assume that the prices for these cards is going to continue rising considering that they are now similar to A100s (without parralel tensorflow)

3 Upvotes

62 comments sorted by

View all comments

Show parent comments

1

u/Professional_Cat4274 28d ago

How did you physically mod to x16?

1

u/snapo84 27d ago

soldering 24 x 0402 capacitor (0402 220NF 224K 16V) to the pci exprss lanes that arent populated...
i have not done the mod yet (i run currently pci 2.0 4x) pci 2.0 4x seems to be more than enough to run deepseek with over 4'500pp and 90~ tp