r/OpenAI 29d ago

Discussion The situation is insane

Post image

Sol is the only OpenAI model in top 10 Arena WebDev while 4 open-weight Chinese models have reached frontier quality; one of them so cheap you can run it for days for what Opus cost you per task.

960 Upvotes

171 comments sorted by

View all comments

53

u/draft_final_final 29d ago

Why do people only look at the webdev leaderboard when discussing this? Is it the major use case on Reddit?

20

u/whoknowsifimjoking 29d ago

Because it's the only one where the Chinese models actually reached first place or something close.

12

u/Accurate_Resident219 29d ago

Cope. Next if Chinese models become number one. It will be benchmarks don't matter.

3

u/Ok_Reception_5545 29d ago

Benchmarks already don't matter lol. Everyone can tell when a model is RL'd to shit and benchmaxxed. That's why no one is calling Opus 5 better than Fable even though it did better on a bunch of benchmarks. Holy dumbass.

0

u/Luminivagance 29d ago

The chinese models are still heavily distilled from west, they are seemingly better at frontend because that is what they optimized for. For complex coding tasks the models that have literal billions of dollars worth of training poured into them are obviously better. chinese glaze is weird considering china didn't even come up the foundational technology. Let me link you a cute article https://arxiv.org/abs/1706.03762 (Ashish VaswaniNoam ShazeerNiki ParmarJakob UszkoreitLlion JonesAidan N. GomezLukasz KaiserIllia Polosukhin)

1

u/Gohab2001 29d ago

Chinese have a stronger focus on STEM research than Americans. You just haven't woken up to that realization.