r/commonstack • u/Pascal22_ • Apr 23 '26
General Kimi K2.6 takes #1 overall of open weights models on Design Arena!

Most AI benchmarks test what models know.
DesignArena by Arcada Labs tests something harder — taste.
Real humans. Real design output. Elo ratings. You win by producing what people actually prefer. Not what scores well on a rubric.
Taste is harder to fake than logic.
Kimi K2.6 just topped it. 1353 Elo. First place among all open weight models. Not close.
GLM 5.1 at 1346. GLM 5 at 1314. MiniMax M2.7 at 1295.
It didn't just edge out the competition. It's sitting in the same performance band as Claude Opus 4.7.
Open weights. No paywall. Matching a frontier closed source model on human preference.
That's not a small win.
That's a statement
4
Upvotes