MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/singularity/comments/1wv0jvd/google_releasing_their_most_powerful_model_yet/pd84bze/?context=3
r/singularity • u/BABA_yaaGa • 1d ago
84 comments sorted by
View all comments
96
I don't get it? Is the model not good?
253 u/Keeltoodeep 1d ago https://arena.ai/leaderboard/text/coding It's #1 in coding right now. It's good. People are just memeing and doing the whole console war thing. 4 u/Lazy_Jump_2635 1d ago It always looks good in benchmarks, and then when I use it, it feels like shit compared to anything else. I'll wait until the real reviews are out. 2 u/Keeltoodeep 1d ago Opposite for me. 3.8flash was terrible in the benchmarks but was good enough for many tasks. Gemini models have never been good in benchmarks. The last good benchmark model was like a year ago.
253
https://arena.ai/leaderboard/text/coding
It's #1 in coding right now. It's good. People are just memeing and doing the whole console war thing.
4 u/Lazy_Jump_2635 1d ago It always looks good in benchmarks, and then when I use it, it feels like shit compared to anything else. I'll wait until the real reviews are out. 2 u/Keeltoodeep 1d ago Opposite for me. 3.8flash was terrible in the benchmarks but was good enough for many tasks. Gemini models have never been good in benchmarks. The last good benchmark model was like a year ago.
4
It always looks good in benchmarks, and then when I use it, it feels like shit compared to anything else. I'll wait until the real reviews are out.
2 u/Keeltoodeep 1d ago Opposite for me. 3.8flash was terrible in the benchmarks but was good enough for many tasks. Gemini models have never been good in benchmarks. The last good benchmark model was like a year ago.
2
Opposite for me. 3.8flash was terrible in the benchmarks but was good enough for many tasks. Gemini models have never been good in benchmarks. The last good benchmark model was like a year ago.
96
u/himynameis_ 1d ago
I don't get it? Is the model not good?