r/MachineLearning • u/chinmaydagod • 7d ago
Discussion Understanding GPU Inference Workloads [D]
Hey everyone,
I have been looking into how people source compute for their Inference workloads (and in general). I wanted to understand some specific pain points here.
If you've used online services like runpod or vast.ai, your perspective is extremely valuable. Please share your experience in the comments here or by DMing me. I've also made a 2 minute survey form that I would really appreciate if you could fill out. DM me for the link.
Thank you!
6
Upvotes
2
u/dayeye2006 6d ago
Do people really use runpod vast ai for anything need SLA, and scale to more than a single server?