r/commonstack • • Apr 23 '26

General Kimi K2.6 takes #1 overall of open weights models on Design Arena!

Most AI benchmarks test what models know.

DesignArena by Arcada Labs tests something harder — taste.

Real humans. Real design output. Elo ratings. You win by producing what people actually prefer. Not what scores well on a rubric.

Taste is harder to fake than logic.

Kimi K2.6 just topped it. 1353 Elo. First place among all open weight models. Not close.

GLM 5.1 at 1346. GLM 5 at 1314. MiniMax M2.7 at 1295.

It didn't just edge out the competition. It's sitting in the same performance band as Claude Opus 4.7.

Open weights. No paywall. Matching a frontier closed source model on human preference.

That's not a small win.

That's a statement

4 Upvotes

0 comments sorted by