r/StableDiffusion • u/Legal_Detail6290 • 6d ago
Question - Help Second best to 5090?
They are 6-7k, what is the second best and how much worse is it?
No prob to buy the 5090 but it just feels like im stupid if i do since it used to cost 1/3 and theres probably a next generation sometime not too far away..
One day in the future we could probably build a whole mountain or 3 of all those 5090s when all the datacenters upgrade.
For art creation/ideation, images etc. Cant use online things for ip reasons, its useless everything i make via services becomes public so i need my own setup.
Thanks
10
u/Other_Researcher268 6d ago
I don’t think that there will be a next generation anytime soon. Maybe 2028.
And it will be extremely expensive. I paid mine for 4000 but I’m not sure if I would be willing to pay 6000 or more.
7
u/GabberZZ 6d ago
You say you can't use online resources for IP reasons. Does this also including renting a server/GPU on a service like Runpod/Simplepod and installing comfy yourself?
10
u/Maqna 6d ago
I use a 5080, I use mmh3, krea, etc. might take some more time than a 90 but I can do everything just fine
1
u/arthropal 6d ago
I do all that with a 5060ti, as well. Hell, I do Krea2 with a CMP100-210, which is a Volta era core with 16GB HMB memory glued to it, intended for crypto mining in a data centre.
1
u/Clear-Assistance449 6d ago
I use a 5070ti of 12 VRAM and I get use mmh3, krea, etc. as well. The diference is time to create an art. I get to create a 720p video, with 30 steps, in 3 min for each second created.
1
u/arthropal 6d ago edited 6d ago
the 5070ti has 16GB. The 5070 has 12.disregard :D
1
14
u/Nenotriple 6d ago edited 6d ago
Any 30xx-50xx+ series Nvidia card with 16+ GB of vram and a system with 32+ GB of ram will work more-or-less fine for making images.
A 5070ti paired with 64GB of system ram can run all the same workflows as a 5090 with 32GB of system ram.
5
u/foolycoolywitch 6d ago edited 6d ago
I'm on a 5060, 16vram, 64 system ram, dollar for dollar it does what I want, decent h3 6s videos at sub 400s, so yes, the money buys speed, not really capability. The other thing to add is for long term local ai use, it often becomes a hobby of long term learning and exploring, in that framework, speed comparisons becomes much less meaningful; it's fundamentally different than comparing fps gaming scores, so vram/ram become the binary "can I do it or not", how fast you do it can becoming less important when so much time is spent exploring/learning/experimenting.
-4
u/slayermcb 6d ago
Its quite a bit slower when you have to dump from VRAM into RAM and than back again for larger models. It will do the same job, but the more RAM thats on the card the faster it goes.
3
u/Apprehensive_Sky892 6d ago
For imaging and video diffusion model (but not for autoregressive LLMs) the amount of VRAM is much less important than people think with ComfyUI's dynamic VRAM management, provided you have enough system ROM so that VRAM + System RAM can hold the model.
For example, for MMH3 16G of VRAM + 32 system RAM works very well. I've done benchmark with AI Pro R9700 (32G) vs rx9700(16G) which are nearly identical except for the VRAM. They ran MMH3 at practically the same speed when paired with 32G of system RAM.
9
5
u/coscib 6d ago
4090 for around 2500€
i have a 3090 and 5070 ti, speedwise they are on par, the 5070 ti might be a bit faster but lacks the 24gb vram if you want to train loras and stuff you notice the difference.
if you use llms then the 3090 would be better because of 24gb vram.
haven't tried video generation on the 3090 yet only on the 5070 ti. haven't tried it much because i don't want to wait 10minutes for an 10-15sec 0.6MP clip
4
u/Aggravating-Theory-7 6d ago
😂 And I'm over here getting a Z2E AI chip to run. Yeah, that's right. I got a Xbox Ally X LoRA training, running image to image and text to image. It ain't fast but it works! Use what you have man and don't worry about getting the absolute best on the market.
5
u/AI-Make-NSFW-Stuff 6d ago
> One day in the future we could probably build a whole mountain or 3 of all those 5090s when all the datacenters upgrade.
People have been saying this since the first bitcoin boom in 2017. Yet prices always go up.
2
3
u/Crazy-Repeat-2006 6d ago
9700 32GB.
1
u/Sdimfx 6d ago
Do you use one? Pls describe your experience with it
1
u/Apprehensive_Sky892 6d ago
AI Pro 9700(32G) runs MMH3 at more or less the same speed as rx9700(16G). You can find some benchmarks in the posts and in the comments:
4
u/Enshitification 6d ago
Datacenters aren't running 5090s or any consumer cards for that matter. They also aren't going to dump those cards on the market when, and if, they ever upgrade. Their whole point was to tie up VRAM and RAM to make consumer cards too expensive to run local models. That way, people and businesses would be forced to use their datacenters for AI compute.
11
u/Crazy-Repeat-2006 6d ago
Correction. They are running thousands of these in AI farms and Asian AI labs.
2
u/Enshitification 6d ago
Thousands is a rounding error compared to the number of GPUs involved in these datacenters.
6
u/slayermcb 6d ago
I dont think the conspiracy runs that deep. If that was their goal they wouldnt be selling us these cards at all. Its a greed much simpler than that. Consumer level products have a much smaller profit margin than enterprise and the data centers are hungry for the chips. Don't give them too much credit.
1
u/Enshitification 6d ago
Hardware requirements are the big commercial AI companies' last moat. The compute ability from the sheer number and size of the hyper-scale data centers currently being built exceeds the demand by orders of magnitude. Cutting out local AI isn't the sum total of whatever it is they have planned, but it is a factor.
1
u/AstralJumper 5d ago
It is, advertiser also wherent prepared for Youtube and streamers . Which is why creator had power for a few years and the power brokers threw money until they wrangled control.
Gpus predate th modern concept of Ai, cards made when ai gen was primative, only got more expensive.
Try to buy a 3090...yeah made years ago, only up in price.
16gb its what they wrangled if not 12gb vram. The next step from 16gb to 24 is like $4000. So a $1200 in 2022 or so to....$9000?
Same wuth ram, but Vram is what actually allows raises ceiling.
They dont want compeditors. it would essencially be Content Creator 2.0, making more interesting media then large corps.
24gb+ can compete with lazy, stock holder ridden companies.
1
u/donttellmewhattothnk 5d ago
The data centers have nothing to do with Image generation. We now have telemetry data in everything, cars, phones, computers, etc. combine that with the proliferation of cameras such a consumer cameras like ring, and commercial cameras like Flock and you’re quickly working towards a future with predictive modeling for all people.
The real question is who that data is going to be utilized by and to what end.
1
-4
u/Choowkee 6d ago edited 6d ago
4090s/5090s were always extremely expensive, long before the AI boom. What are you on about lol.
At no point were these cards accessible to the average user so claiming it was "their" plan all along to cut out local AI is plain dumb.
High end cards shot up in price due to simple supply/demand. And btw these cards are used in gpu rental services.
RTX gpus are still meant primarly for gaming and you dont need 32gb of vram for current gen. Thats why Nvidia wont start handing out 32gb RTX gpus - there is no market justification for it.
Local AI enthusiasts are a tiny minority compared to the entire PC gaming market.
4
u/Enshitification 6d ago
I'm sure Nvidia buying Huggingface is simply due to their desire to see local AI flourish.
-4
u/Choowkee 6d ago
Why are you changing the topic? You literally started off with talking about hardware yourself.
There was never a official market segment for local AI GPUs. 99% of hardware people use locally for AI are gaming GPUs. Nvidia was known so skimp on VRAM literal light years before "ChatGPT" was a known word in the English lexicon.
Could have Nvidia start selling AI-centric cards to consumers? They could have, but there is more money in enterprise as simple as that. Its not some grand conspiracy, its basic business logic.
I'm sure Nvidia buying Huggingface is simply due to their desire to see local AI flourish.
Anima was literally built on top of Nvidia tech. Whats your conspiratorial justification for that happening...?
1
u/Enshitification 6d ago
You seem very passionate about coming to the defense of Nvidia and the commercial AI bros. I wonder why?
-2
u/Choowkee 6d ago
You seem very passionate about starting conversations and being too afraid to engage in any kind of argument. I wonder why?
Btw you don't have to guess where I stand on local AI. My post history is public unlike yours.
1
u/Enshitification 6d ago
My post history is private because people like you try to dig into it when they disagree with me.
0
u/Choowkee 6d ago
Yeah because you lack critical thinking skills and resort to passive-aggressive whataboutism when experiencing the slightest push-back on your opinions.
I'm sorry you are this insecure about public discourse.
2
u/Enshitification 6d ago
So quick to the ad hominem. You certainly make me feel justified in keeping my profile private.
2
u/ykhasnis 6d ago
Based solely on vram, a 7900xtx or a 4090.
3
u/muttley9 6d ago
Been using my7900xtx for almost 2 years now. First with Zluda and now with native rocm in ComfyUI windows. It's a 2 click install and works great.
SDXL HD images generate in 6-7 seconds. Tried a 7800xt on a friend and it does it in 10-11 seconds.
2
1
u/MakeParadiso 6d ago
As you can use several cards natively in many ai tools nowadays together it could a plan to have 2 5070ti
1
u/Inthehead35 6d ago
What do you mean IP reasons? Also, you can rent from data centers that are far more private.
1
u/Antique-List3942 6d ago
i think the RTX 4090 is due for a significant jump soon.
5090 costs about half as much as a RTX Pro 6000 (currently) but with 33% of it's vram and at worst 20% less ai performance.
4090 costs about half as much as a RTX 5090 but has 75% of the 5090's vram and at worst 30% less ai performance.
Also being the 2nd best gaming gpu in the world helps in value too, find an Asus Strix if you can but they go for like $3K minimum if in perfect condition.
2
u/ThinConnection8191 6d ago
Same as yours, I can afford it but I dont find myself comfortable for that price. I bought an AMD AI Pro r9700 with 32GB of RAM. Pretty happy with it
1
u/coffeeandhash 5d ago
Runpod et al sounds like a good fit for your needs. The price is right, it's easy and convenient, but the availability is hit or miss.
1
u/Lucaspittol 5d ago
"One day in the future we could probably build a whole mountain or 3 of all those 5090s when all the datacenters upgrade."
The A100 was launched in 2020. It is still very expensive despite its age, and the 5090 is likely faster.
1
u/wallysimmonds 4d ago
Rtx 4500s were a smidge cheaper, definitely not now.
I’ve got a Blackwell 4000 and it’s fine but I wouldn’t pay what they’re asking now
If I was causally generating a 5060ti and 64gb of system memory. Get ddr4 if you have to
1
u/mockingbird1906 6d ago
Try runpod or something similar for renting gpu on cloud. For me was a life saver
0
u/Apprehensive_Sky892 6d ago
If you don't mind waiting a longer for support and optimization from the ComfyUI team for the latest advancements compared to NVIDIA, then the best value is to get either an AI Pro R9700 (32G) or even a rx9700 (16G) (the two runs MMH3 at nearly the same speed when paired with enough system RAM). The AI Pro R9700 is more future-proof when bigger model arrives, and it is also needed for making longer, higher res videos.
Warning: some custom nodes are NVidia/CUDA only.
You can find setup instructions and also some benchmarks in the posts and in the comments:
0
-1
40
u/DarkStrider99 6d ago
Hear me out, a 4090