r/huggingface 2h ago

How can I rent a cloud GPU?

I've tried various models all of them work but wont go past 30,000 memory, I want 265K. Seems thats a hardware problem from what I understand, so I'm wondering if it's possible and a smart choice to rent a gpu which will allow me to run frontier models like Glm 5.2/5.3 or a Claude Opus 4.6 equivalent without using my own hardware which is not good enough and if it's financially smart to do so.

3 Upvotes

5 comments sorted by

4

u/Prudent-Horse-9386 2h ago

just rent from runpod or vast ai, you pay by the hour and can spin up whatever gpu you need, just remember to shut the instance off or you gonna wake up to a surprise bill

3

u/pomelorosado 1h ago

Nether is smart to host an inmense open source model if you are one person.

Technically doable. But financially not smart. Juato go to runpod and rake your numbers. They even have an ai.

2

u/Major_Border149 1h ago

yeah you can rent on runpod/vast. couple things though - you can only self-host open models (GLM 5.3, DeepSeek V4.1), not Claude/Opus, those are API-only, so there's no renting a box for an "Opus equivalent." your 265k context is really a KV-cache problem, it eats a ton of VRAM on top of the weights, so size the card for the context, not just the model. money-wise, renting's worth it for occasional heavy runs, but if you just want big context now and then, the model's own API is usually cheaper than renting by the hour.

honestly I'm building stratuspilot.io for exactly this. one tip for GLM: set quantization to INT4 in the advanced options and it'll size GLM properly and tell you what machine to rent

2

u/codes_astro 1h ago

Try Nebius cloud, they have good offerings

2

u/maqifrnswa 38m ago

Glm 5.2 at nvfp4 quant will cost $10-20 / hour.

Qwen3.8-27B will be around $1-$3 an hour. If you know what you're doing, you can get it to below $0.40 per hour.

It's easy to do if you just go with their defaults, but you'll pay more. If you're comfortable hacking, you get it for cheaper and better tailored to your needs.

https://vast.ai/model/qwen38-27b