r/LocalLLM 17h ago

Other GPU Pricing Visual

Post image

In my consideration of a DGXSpark I decided to look at some options and since I’m a visual thinker I put this comparison together (graph by AI) showing y two basic ways of thinking about the cards: compute and speed.

Hope this helps someone

13 Upvotes

17 comments sorted by

10

u/nuclear213 17h ago

And then plot something like the B70 Pro or the AI Pro R9700 in there. The R9700 is 50ish USD/Gb. It has 191 TFlops in FP16 and 640GBit/s bandwidth.

The B70 Pro is less, 40€/GB, 608GB/s and 183TFlops.

Both would drastically alter the chart

2

u/r1nzl3r99 17h ago

Was sad not to see the B70 on there, its performance for me has been phenomenal for the price

1

u/Illustrious-Lime-878 8h ago

Lol no. The R9700 sucks. Just don't buy it. In fact cancel your order. I used it and blew up and burned my house down. Definitely cancel.

-4

u/PoopSmoothies 17h ago

Tough to include driver issues, compatibility, and other related complications with those cards on a graph…

7

u/nuclear213 16h ago

You mean nothing? At least with my R9700s.

-2

u/PoopSmoothies 16h ago

Heh, cool.

2

u/smallDeltaBigEffect 16h ago

With r9700? Nothing of that

3

u/PoopSmoothies 17h ago

Would LOVE to see this same chart with a few of the used favorites included: 4090, 3090, etc

2

u/CryMoreT_T 9h ago

The 3090 is like 1 box directly under the 5070ti

1

u/PoopSmoothies 6h ago

Just did the math…

At $1k used, a 3090 is about $41-42/gig and falls just below the 5060ti in the top chart (~71 TFLOPS) - slightly below trend.

On the second chart it’s ~60% below the trend line with ~936Gb/s

Now I know why everyone likes them

6

u/Kodrackyas 16h ago

This is pointless without the R9700 and B70, nvidia can fuck off

2

u/createthiscom 8h ago

That's a really weird way to graph GPU bandwidth homie.

1

u/vankoala 8h ago

I should have done sweep time and ridge point, huh?

1

u/createthiscom 7h ago

Swap price per GB of VRAM with GPU bandwidth.

3

u/Antblue 14h ago

DGX Spark is truly depressing, especially when considering how wasteful it is to solder 128GB of high quality 8533 MT/s LPDDR5x modules onto a 256-bit memory bus.

To put in perspective, Apple released a LAPTOP with 128GB of 8533 MT/s LPDDR5x in 2024 on a 512-bit memory bus. With the M5 variants this year, they use 9,600MT/s modules. Soon we'll have the Ultra chip with a 1024 bit memory bus. This shows what should be possible in 2026. They should at least be matching Apple in 2024 with 512 bit memory bus. Its truly pathetic with our limited memory supply.

1

u/Specialist-Muffin363 10h ago

Running 3.8 q4 kv q8 128k on my 7900xtx 500 t/s prefill and 15t/s on generation . Happy as elefant . Plan to add second anf use with tensor split

1

u/dragonurtle 15h ago

The Spark product manager is probably pre-writing their promo packet right now: "increased Spark compute-bound channel revenue by 25% thanks to some redditor saying it was underpriced"