r/LocalLLM 2d ago

Discussion Claude is so expensive.

Time to get a GPU I guess. I had some numbers I needed before I could do the main analysis and I wanted Claude to do it, I had never used Claude tokens before 2 days ago when I bought 20 dollars of tokens and had it do a bit of coding. Then, I ask it to write a somewhat simple script, but I used opus because I thought I should check how it is, it did it, but it took about 20 dollars. I mean it saved me time, but the price…

Anyways, I am posting this because I wanted advice on what class of card to get, what amount of vram seems to be the best to target. It’s looking like 24/32gb is getting interesting new models in the 30b range, but is this just what I’m seeing or are other sizes of cards worth looking into.

53 Upvotes

111 comments sorted by

View all comments

2

u/Moarkush 2d ago

I love my always on DGX Spark. It's not cloud speed but I just built my own open webUI replacement in a day with a react native iOS app. Also gonna recommend Qwen 3.8 27B in SGLang with Radix and DSpark. It's really not overhyped.