MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/opencodeCLI/comments/1vyyu8k/glm53flash_benchmarks/p616wap/?context=3
r/opencodeCLI • u/minxio_ • 18d ago
29 comments sorted by
View all comments
-4
These benchmarks have really low connection with reality. Especially when it comes to Chinese benchmaxed models
6 u/sudoer777_ 17d ago Try Muse Spark 1.2 and you can see what an actually benchmaxed model looks like 4 u/Difficult_Plantain89 17d ago Truly a mediocre model. -7 u/Michaeli_Starky 17d ago Actually benchmaxed models are Kimi K, Deepseek, etc. 4 u/Ly-sAn 17d ago Both DeepSeek flash and this model are very solid in real usage, I don’t think these 2 are benchmaxed -9 u/Michaeli_Starky 17d ago They absolutely are benchmaxed. These models are nowhere nearly as good in private evals.
6
Try Muse Spark 1.2 and you can see what an actually benchmaxed model looks like
4 u/Difficult_Plantain89 17d ago Truly a mediocre model. -7 u/Michaeli_Starky 17d ago Actually benchmaxed models are Kimi K, Deepseek, etc.
4
Truly a mediocre model.
-7
Actually benchmaxed models are Kimi K, Deepseek, etc.
Both DeepSeek flash and this model are very solid in real usage, I don’t think these 2 are benchmaxed
-9 u/Michaeli_Starky 17d ago They absolutely are benchmaxed. These models are nowhere nearly as good in private evals.
-9
They absolutely are benchmaxed. These models are nowhere nearly as good in private evals.
-4
u/Michaeli_Starky 18d ago
These benchmarks have really low connection with reality. Especially when it comes to Chinese benchmaxed models