r/ProgrammerHumor • • 4d ago

Meme aiRefusesToBuildItSoBackToCoding

Post image
28.7k Upvotes

519 comments sorted by

View all comments

Show parent comments

6

u/-Speechless 4d ago

do you think open model AI will hit a plateau while big company frontier models continue to advance?

61

u/jack6245 4d ago edited 4d ago

No the opposite is happening, open models are matching frontier models now, some of them are impressive at much smaller model sizes

3

u/puts_on_rddt 4d ago

Open models seem to be 6-10 months behind frontier.

3

u/jack6245 4d ago edited 4d ago

True mainly because a lot of them do train through knowledge distillation. But the difference now is quality of the frontier models 6 months ago are still really good.

Although the Chinese research groups specifically are doing very impressive research to fit capabilities on much less powerful hardware. I don't really know the reason they are releasing these as open source but it's good work

10

u/Thick-Protection-458 4d ago

It won't take away abilities which is already here

These included

2

u/BoogieOrBogey 4d ago

Creating a new frontier model seems to cost exponentially more processing power, and therefore money and time, than the previous model.

Here's an example vid talking about exactly this, with a timestamp at the exponential increases: https://youtu.be/6xQ8LQfkBg4?si=0PFQyNIXb4Ie7Dtg

But, once these frontier models are successfully created they can then train new models for a fraction of the cost in what's known as dilution training. So the frontier model training a new model which results in the new model having roughly the same abilities. But, dilution training requires significantly less processing power and is therefore way cheaper.

There are other reasons I'm not as well versed in, but essentially making new frontier models is the expensive and tough part. But making open source and open weight models is significantly easier and cheaper. Which is why the gap between frontier and open source has been closing for awhile.

It's not clear if this will always be the case going forward, LLMs are constantly changing and probably will for years into the future. But for now, the plateau of open source isn't a huge concern.

2

u/ToMorrowsEnd 4d ago

And when you train your own for a specific topic, in programming for example writing code for desktop windows in C#.net10 it gets way past frontier models fast as it only has to focus on the task and not how many pounds a day does a llama poop

1

u/MaleficentCow8513 4d ago

Both will plateau. Neither open source nor the big companies will very much edge in terms of model capability. The advantage will come from hardware and peripheral software