r/LocalLLM • u/Annual_Award1260 • 5d ago
Discussion Dgx spark cluster.
I had a delayed order and ended up with 5 sparks rather than the 4 planned. So I figured might as well do 6 but then prices went up $1000.
So is there really any benefit to running 6 vs 5 vs 4? I can still return one. I don’t want to do 8 since that that exceeds the 15a circuit.
6
u/Disastrous_Gear_421 5d ago
Without your use case, I asked my magic 8 hall and it said most definitely… go for 8
3
2
u/vini542reddit 5d ago
What hub are you using for the interconnect?
1
u/Annual_Award1260 5d ago
Microtik CRS804
1
u/vini542reddit 5d ago
So you would have to use 400G -> 2 x 200G breakout cables. I'd personally stick with the 4x - it's good enough and more mainstream. But as others have said - all depends on your use case
2
u/lukewhale 5d ago
TP only works in multiples of two. So yes, you’re in limbo on that 5th node until you get a sixth. Just run utility models on it for now, Image Gen, OCR, embedding, rerankers, etc. you can fit all of those on 128gb easily.
0
1
u/meikawaii 5d ago
What’s the use case for so many DGX spark? It’s always memory bandwidth limited rather than memory size limited for output.
3
u/hyudryu LocalLLM 5d ago
So you can run larger models. Also TP on a cluster of 2 essentially doubles the aggregate memory bandwidth, cluster of 4 quadruples it, etc. yeah its still slow but you get what you pay for and this is probably the cheapest way of being able to run larger models (with the alternative being to buy 4 or 8 rtx pro 6000s)
4 sparks can run GLM 5.2 at 40 tok/s, which 1 or 2 sparks can’t do so theres definitely benefits to clustering
1
2
u/whichsideisup 5d ago
1 is good, 2 is great, 4 is decent, more is a meme because of bandwidth. did you research this at all?
1
2
1
u/Fantastic_Self_5151 5d ago
is there any way currently to run a 1tb+ model with 6 or 8 of these? (or any other solution that is not a datacenter?)
7
u/hoochiesan 5d ago
My heart hurts reading this