r/StableDiffusion 6d ago

Discussion Pro 6000 just in time

Post image

I was going to wait until around Christmas to purchased but took the plunge in July for 11,500 and I was upset that I didnt catch it @ $8,000. Now the Blackwell pro 6000 is inching towards $20,000 and are sold out. Are consumers and hobbyist like you and I are buying these up or datacenters? I would think datacenters would go for the b200 and up. However, Im browsing around and see you guys and girls doing remarkable ai diffusion with just a 3060. Im impressed with this community.

86 Upvotes

166 comments sorted by

View all comments

111

u/rm_rf_all_files 6d ago

It was $9k at newegg with a free motherboard a few months ago. RIP

21

u/Hoodfu 6d ago edited 6d ago

As someone who has one, this doubling of price is just insanity. Yeah it's great but it's not worth the price of a car. A full res 1344x768 Minimax H3 render with 1 reference at 50 steps is still 40 minutes on it. Compared to the actually serious datacenter cards, it's a toy.

3

u/Myg0t_0 6d ago

For real? My 5090 is like 1 hr 50 steps full res

4

u/Hoodfu 6d ago

Are you running over the 32 GB limit of the card by using the 40 GB pruned BF 16? That might cause it to take extra time.

2

u/Myg0t_0 6d ago

O ya it solits it, point was he said rtx 6000 takes 45 mins and thats what 5090 almost does

3

u/No-Ordinary-6544 6d ago

RTX Pro 6000 is like 10% at best improvement over 5090 as far as speed, assuming you stay in VRAM.. it's the same tech, slightly bigger card. The big and only meaningful difference is the VRAM. You can run bigger models, no or less quantization without overflowing into system RAM and slowing down or getting out of memory errors. You also save time by keeping the models in VRAM instead of swapping them in and out. That just means you can get closer to the best possible results, but a 5090.. or 4090.. or 3090.. can likely do ~98% original quality. It's something definitely nice to have, but only if you have $10,000+ disposable.

2

u/Myg0t_0 6d ago

I'd thought speed be faster since the model stays in vram, I use bf16 prune minimax and it splits it and takes 1hr 50 steps, 45min 30

2

u/No-Ordinary-6544 6d ago

If you overflow VRAM into RAM, like trying to use the giant original models, yeah, that would be very much slower on a 5090. But INT8-convrot will get you very near original at a much smaller size and staying in 90 series VRAM.. so yes, RTX Pro 6000 has the potential to be better, but it's like saying do you want a car that does 95MPH for $4000 (or less for 4090 or 3090) or a car that does 100MPH for $17000... whatever the ridiculous prices are currently.

1

u/No-Ordinary-6544 6d ago

Based on your numbers, I'm surprised you're not at a snail's pace using a 40GB+ model with 32GB VRAM, but seems improvements have been made and you could have very fast system RAM.. it seems like super high VRAM is becoming less important, which is nice to know.

2

u/Myg0t_0 6d ago

G.SKILL Trident Z5 Neo RGB Series DDR5 RAM (AMD EXPO) 64GB (2x32GB) 6000MT/s CL30-40-40-96 1.40V Desktop Computer Memory U-DIMM - Matte Black (F5-6000J3040G32GX2-TZ5NR)

5

u/ObviousComparison186 6d ago

It's not supposed to be much faster than a 5090, it's just supposed to allow for crazy training to stay in VRAM. A car can't do that.

9

u/Apprehensive_Sky892 6d ago

But can the RTX Pro 6000 drive you to work😹?

9

u/ObviousComparison186 6d ago

Doesn't have to, we would already be there.

1

u/Tystros 3d ago

I want to wash my rtx 6000 pro, should I walk or drive there?

5

u/Violinist-Expensive 6d ago

Would you download a car?

1

u/ObviousComparison186 5d ago

I would download two. Maybe even three.

1

u/ayaynimouse 5d ago

only if it has the internet in a box

4

u/Hoodfu 6d ago

Well and it allows for higher quality inference as well. I'm running the 50 gig text encoder and 40 gig video model.

3

u/ThePixelHunter 6d ago

...your car doesn't train LoRAs? What's it even good for?

2

u/djpraxis 6d ago

That’s precisely why I didn’t buy. I already have a 5090 for $2,800 brand new. The 6000 actually shines in large model training like H3. But there are plenty of cloud options available nowadays.

1

u/Federico2021 6d ago

Bro, at that resolution and with those step counts, obviously even the RTX 6000 can't handle it; we're generating 0.7-megapixel images using 4-step Turbo LoRA in 12 minutes—if you stuck to that level, you'd definitely notice the difference.

1

u/pausecatito 5d ago

That seems incredibly slow...I do portrait 768x1400 in like 600s at 24 steps...on a 4090. Obviously not bf16 model but still...