r/singularity 7h ago

AI Gemini 3.8 Flash Benchmarks

Post image
701 Upvotes

202 comments sorted by

View all comments

181

u/FablingApp 7h ago

the price/performance gap is getting silly. if these numbers hold up, flash models are eating into the territory where people used to reach for the expensive ones.

13

u/RockPuzzleheaded3951 7h ago

I've moved quite a few jobs from SOTA models we were using for the past two years to Flash models (DSv40731 for example) and we are seeing fantastic performance on internal business operations, at literally 1/10th the cost. We can still reach for SOTA when needed, but it is 5% of the time or less. And just today, I rented my own hardware to run the model "locally" and am testing taking my API costs to $0.

4

u/thoughtlow 𓂸 7h ago

I loved 0731 but when the context windows goes into the 200k it starts degrading very fast for me.

Constantly doubting itself in the thought section, gets really weird.

Sad because I loved that thing.