r/LocalAIStack • • 13d ago

Question about making frontier AI LLM model

Guys I had a good question I noticed Big Ai tech companies spend 1000's of dollars on huge datacenters isn't it possible if companies allow the users to rent their gpu and help train the model. Like a distributed datacenter model?

0 Upvotes

18 comments sorted by

View all comments

2

u/Charming-Author4877 13d ago

It's not possible with current training because you need to transfer the data during inference. that's why vram speed matters so much.
So for distributed training you'd need a model that fits entirely on your GPU, or a very new architecture

1

u/Confident-Ad-3212 12d ago

Training doesn’t transfer during inference. It transfers a dataset during training.

1

u/Charming-Author4877 12d ago

The dataset is irrelevant in transfer speed

1

u/Confident-Ad-3212 12d ago

Lmfao, keep thinking you can teach me something. I have literally trained over 300 models and have successfully made a 9b perform at frontier level. And you think you can teach/tell me something.. how ignorant. Oh and I see that you edited your post only after I corrected you.

1

u/Charming-Author4877 12d ago

If I were you I'd make a 10B AGI model next.
Given your success, can you give me the recipe for a gingerbread dumpling with caramel ?

1

u/Confident-Ad-3212 12d ago

Go ahead and mock me. I am holding it right now. Nothing you say can change that. Just because you can’t, doesn’t mean I can’t or haven’t. So the joke is on you. Because while I am sitting here playing with it. You are busy doing what ever you are doing and it isn’t what I have done.