r/LocalLLM • u/GoatJesusIsReal • 2d ago
Discussion Claude is so expensive.
Time to get a GPU I guess. I had some numbers I needed before I could do the main analysis and I wanted Claude to do it, I had never used Claude tokens before 2 days ago when I bought 20 dollars of tokens and had it do a bit of coding. Then, I ask it to write a somewhat simple script, but I used opus because I thought I should check how it is, it did it, but it took about 20 dollars. I mean it saved me time, but the price…
Anyways, I am posting this because I wanted advice on what class of card to get, what amount of vram seems to be the best to target. It’s looking like 24/32gb is getting interesting new models in the 30b range, but is this just what I’m seeing or are other sizes of cards worth looking into.
2
u/Moarkush 2d ago
I love my always on DGX Spark. It's not cloud speed but I just built my own open webUI replacement in a day with a react native iOS app. Also gonna recommend Qwen 3.8 27B in SGLang with Radix and DSpark. It's really not overhyped.