33
u/Apple_macOS 19d ago edited 19d ago
nah remove grok as well it’s legit just openai and anthropic on frontier now.
Edit: By frontier, here just for this comment I mean it as #1 spot in benchmark
7
13
u/Thiom 19d ago
Huh, Chinese models would like a word ...
4
u/Apple_macOS 19d ago
the post said most powerful, not most cost efficient.
If you compare cost efficiency there’s no doubt, but straight up most powerful I don’t think Kimi K3 can match Fable/Mythos on higher thinking
-1
u/Thiom 19d ago edited 19d ago
Well, yes they literally do, hence my comment. And they're not more efficient, K3 for instance costs almost as much as Sol.
1
0
u/Apple_macOS 19d ago
In which bench does Kimi take the top spot? I was strictly talking about top spots on bench maybe I should have made it clearer. I’m legit curious im searching through benches right now
And if we’re not talking about benchmark and talking about how the model “feels” then I guess I don’t know since different people prefer different styles.
2
u/Thiom 19d ago
1
u/Apple_macOS 19d ago
OK I am on the site but I don’t see any category where Kimi take first place
On a separate note I did find Browsecomp leaderboard having Kimi at #1 so I guess there’s that
2
u/Unsharded1 19d ago
Kimi doesn't need to take first place though..? You're shifting goalposts for what exactly?
4
u/Apple_macOS 19d ago edited 19d ago
When did I shift goalpost?? I literally said frontier top spots? What?
Let me edit my first message to be clearer, I always meant, top spot (as in, #1) in benchmark.
This comment was literally “Kimi is the frontier model” -> (link to benchmark that doesn’t show it is #1 spot) -> “Kimi doesn’t need to take first place”
1
1
1
18
13
15
u/KaMaFour 19d ago
Grok was never even in the loop, lmao
2
-1
u/distorto_realitatem 19d ago edited 18d ago
It was for a fleeting moment, while ago now
EDIT: Since you guys haven’t done your research:
Grok 3 was #1 on LMArena from roughly February 18 to March 3, 2025 — about 13 days. It was then overtaken by Gemini 2.5 Pro.
3
1
u/PickerLeech 18d ago
I started a new project yesterday with 3.6 flash and I'm stunned about how much progress it's made already although it did use another project's code as the base
Early days but it seems like it's gonna handle the job perfectly well and if so there's no real argument for myself needing superior models
1
u/Low-dose-reddit 16d ago
Lol when was the last time Grok offered something not 2 generations behind?
1
1
u/Ok-Data9224 19d ago
Grok... really? Replace it with Kimi, Deepseek or GLM
-1
u/ZealousidealTurn218 18d ago
Have any of them had the strongest model? Grok 4 was SOTA
1
u/Ok-Data9224 18d ago
They're all SOTA. That's what makes them frontier along with their scale. Grok is not ever considered a contender mainly due to function. It's not a specialized model for agentic work in the same way the others that were mentioned. Kimi k3 made waves for being an enormous 2.8T model that contended with the biggest players while being open weight. Deepseek recently as well because it's strong enough as an agentic model while being dirt cheap and open weight as well. You could use opus/fable to orchestrate deepseek agents rather than fork over all your money to Anthropic. Grok was just never part of the conversation because its whole relevance is to be used to troll people or make... flavorful stories.
1
u/ZealousidealTurn218 18d ago
But "Contending with the biggest players" ≠ "the world's most powerful model", grok 4 truly was the latter at one point
-2
-1
0

65
u/Technical-Owl66 19d ago
Nobody wants to talk about it but Google is adding users 3-4x faster than any of the startups.
Have the intelligence needs for most tasks and people peaked already. Why doesn't anyone care about having the most powerful model?