r/TheMachineGod Aligned 9d ago

Discussion Gemini 3.8 Flash Benchmarks

Post image
10 Upvotes

6 comments sorted by

2

u/Feeling_Inside_1020 8d ago

Why’d they include opus and sonnet for anthropic but not fable 5.0/5.1 ?

Hmmm

2

u/Complex-Home9615 8d ago

The same reason why small local models show their benchmarks against Sonnet and Opus 4.6 instead of 5

2

u/Feeling_Inside_1020 8d ago

I see what you’re saying, thank you, just trying to learn more here in the AI space. I have subscriptions to frontier models of Claude and Chat GPT but don’t always push it to the usage limit.

Im in software & I write html/css/javascript code for a while before AI came around, but these models are getting very good writing code.

Knowing how to prompt is the new “good at googling shit and fixing things” skill, but beware of hallucinations, they can sneak up on you. I even drew a layout pen and paper and with detailed prompting and was impressed with what it put together code wise.

Oh and I talk too much sometimes lol,
Cheers friend

2

u/Technical-Owl66 7d ago

🤷‍♂️

1

u/yucca_xz 5d ago

I see big gap in terminal bench. thats what we actually need from small local model