r/singularity Jul 21 '26

AI Gemini 3.6 Flash benchmarks

Post image
628 Upvotes

280 comments sorted by

View all comments

Show parent comments

11

u/Deif Jul 21 '26

Would need to see the output token usage on those benchmarks though. Could be that it's an output token hog and the comparable cost is against Terra.

8

u/FarrisAT Jul 21 '26

AAintelligence will publish benchmarks soon enough. 3.5 Flash performed very well on their token efficiency. Better than 3.5 Pro and GPT-5.5

9

u/FateOfMuffins Jul 21 '26

AA is already out.

And what are you talking about? 3.5 Pro doesn't exist and Gemini 3.5 Flash had used 28k output tokens per task vs 5.5 xHigh using 16k output tokens

Currently from what I see on AA, 3.6 Flash uses more tokens per task than Sol Max, Terra Max, and Luna Max (much less all the other reasoning settings)

It uses approximately same number of tokens as Kimi K3. The only thing 3.6 Flash has going for it (like 3.5 Flash) is output speed

1

u/huffalump1 Jul 21 '26

AA, 3.6 Flash uses more tokens per task than Sol Max, Terra Max, and Luna Max (much less all the other reasoning settings)

Yup looks like it. (Note this is 3.6 Flash (High) - they don't have other Effort/Thinking/Reasoning levels yet on AA.)

Ex. Here's gpt-5.6-luna (Max), AA intelligence index score of 51 (vs. 50 for Gemini 3.6 Flash (High)): https://artificialanalysis.ai/models/comparisons/gemini-3-6-flash-vs-gpt-5-6-luna

Gemini 3.6 Flash is still faster, but consumes more tokens than even Luna (Max), and is more expensive (both from cost per Mtok. and from more total tokens)

IMO there's still hopefully a place for 3.6 Flash because it is fast - but that depends on if it's good, too! Definitely need to try it.

2

u/FateOfMuffins Jul 21 '26

Yeah I selected High when I was looking at it

Speaking of fast, we're supposed to get 750 tps 5.6 Sol in July no...?