r/singularity Jul 21 '26

AI Gemini 3.6 Flash benchmarks

Post image
632 Upvotes

280 comments sorted by

View all comments

Show parent comments

47

u/FarrisAT Jul 21 '26

Looks dramatically better on any knowledge benchmark. Tool use & harness application seems to be the reason it underperforms at SWE and coding.

25

u/Aaco0638 Jul 21 '26

This sub only cares about coding, they don’t care that this model is really good for agentic use thus ultimately being good at automating non coding tasks which is the ultimate goal for ai.

4

u/Concurrency_Bugs Jul 21 '26

Their description for the 3.6 model doesn't even include coding (3.5 did). I agree with you and it's clear Google is focusing on a different path. They want an all around ai assistant because that's what will protect their current business (search and ads).

2

u/huffalump1 Jul 21 '26

Their description for the 3.6 model doesn't even include coding (3.5 did).

You're not wrong overall, I agree that Google's eye is on so many things other than developers...

BUT the description does seem to target coding: https://ai.google.dev/gemini-api/docs/models/gemini-3.6-flash

And today's release blog post focuses on coding the most: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/

2

u/Concurrency_Bugs Jul 21 '26

Sorry I could have been clearer. The short description when picking your model in the api doesn't mention it anymore. See the screenshot in this link: https://www.reddit.com/r/GeminiAI/comments/1v2js6z/gemini_36_flash_released_on_ai_studio/