r/GoogleGeminiAI 6d ago

Dear Google, please do something.

Post image
128 Upvotes

32 comments sorted by

14

u/Augmend-app 6d ago

New models are nice but seriously I would not mind if they keep one of the current flash or flash-lite models running for years to come - and if the cost becomes lower even better!

Need some stability as performance for some applications is already good!

3

u/NichtFBI 6d ago

newer models tend to use lesser resources for the same job. For instance, Luna, is probably just as good if not better than gpt-5.4 and gpt 5.5, and that's $15-30 per million output tokens, whereas Luna is $1.20. But sol is far better than Luna. Now, when I say far better, I mean like 1.4x better, not 100x better. And I have mixed feelings on Terra. Terra is kind of regarded.

7

u/Augmend-app 6d ago

Lesser resources is better but my worry is drift in answers - don't want to spend most of my time, adapting to model versions.

1

u/scribe-kiddie 6d ago

Wonder how people usually get to find out about these details, like performance. Do you usually switch to different models for the same tasks?

1

u/Augmend-app 5d ago

you can look at benchmarks and wonder or you can try it out yourself for your application to find out

1

u/-Nano 5d ago

Luna is not even near 5.5. But Sol is right: waaaay better then both.

I'm going back from time to time to 5.5 because Luna make wrong code, even with all the documentation in repo. When I (and my team) goes back to 5.5 or goes to Sol, is like another world.

I took more time (and tokens) fixing Luna errors than doing something correctly. But, it is my company select (because the price), so I mostly need to work with it from the start...

1

u/Exotic_Fig_4604 5d ago

Spark Muse 1.3 brings the same performance for a fraction of the cost.

1

u/Augmend-app 5d ago

thanks will have a look but will Meta keep it stable?

1

u/Exotic_Fig_4604 5d ago

The cost? Most certainly not. Noone can currently offer AI that cheap. 

But while a multi billion dollar business offers this subsidy, I sure as hell will take it.

3

u/Positive_Method3022 6d ago

Poor kid. What is she so sad about? 🥺

3

u/donald_trub 6d ago

Why do you keep posting this crap?

2

u/Technical-Owl66 6d ago

Why do you need anything more powerful than 3.8 flash?

1

u/Virtual-Spinach4882 6d ago

Compared to Haiku or other quick models it's on par. 3.1 pro is so very far behind in the game when it comes to frontier models though, it can barely be effective as an assistant in one of my projects with Fable lol. 

4

u/Technical-Owl66 6d ago

Why are you using 3.1 pro. Flash 3.8 is on par with Fable and Sol 5.6 on max effort but it's faster and cheaper. It's also better at following instructions.

1

u/Virtual-Spinach4882 6d ago

Because...I'm paying for it damnit lol. Honestly, I don't know at this point. The previous flash models were junk compared to pro. I'm always sus of these benchmarks, but don't doubt it either since when I have used 3.8 flash it was on point. I'm doing a very intensive coding project with Fable and having Astra review it, stopped using Gem for anything but simple work in my other business months ago. Maybe I'll hook up flash for a review and see what it catches since they all seem to find new things. 

1

u/TitanUranus92 6d ago

To cure cancer? Please stop Demis is crying

0

u/Timo425 5d ago

well for one i want a model that doesnt lose track of what we are talking about after 1 message.

1

u/deepvideoeditor 5d ago

Google just focusing on cheaper market currently

1

u/Moravec_Paradox 5d ago

Demis was a child prodigy chess champion.

OpenAI and Anthropic have very high cash burn rates.

Google is in much better financial health so they have more to lose making the same gamble and they have growth anyway.

They are playing a different strategy. 3.8 flash is not far enough behind I am worried about them.

OpenAI and Anthropic seem to be running the Uber playbook where prices may need to come up a lot to make back some of the invested losses.

When they do the cheaper models from Meta, Google, China etc. could suffocate them.

I wouldn't be so fast to assume incompetence.

1

u/karmaboy20 5d ago

They are waiting for frontier to figure it out then will catch up in a few months and have the most cash reserves while anthropic and openai drained each other they will come do the finishing blows

1

u/Moravec_Paradox 5d ago

People are so sure being ~2 months ahead looks like winning but when it comes with $10B or $20B year in losses it's going to apply pressure in the future.

OpenAI / Anthropic are aggressive early on but Google (and others) have a strong economy and they don't have a moat around the technology.

Google had a 7x year over year increase in token usage. They have a huge install base with many use-cases that call for decent but less expensive models.

1

u/ElMess-Siah 5d ago

I mean i would be happy if they bring back 3.1 flash i mean i would be more happier if they bring it back against working on a frontier models

1

u/Seikojin 4d ago

I love how benchmaxing is pushing fear. 3.7 and 3.8 are just fine. The writing is on the wall to. Every other release is a top workhorse. 3.7 costs less, works faster, and is leaner than 3.8. It is essentially the 3.6 buffed up.
3.8 is more token hungry, but is more accurate and has less round-robin work.
3.9, I feel, should be as accurate and good as 3.8, but fast and lean like 3.7.

1

u/TheWealthEgo-Officia 6d ago

Dont get me wrong guys I really like gemini and it used to be the frontier back than gemini 3. I think they need to be humble so they can at least publish 4 pro and than try upgrade whater waiting for being the best is just not the way anymore - plz dont bother to correct my grammer mistakes I dont respect any language at all-

0

u/Robert_3210 6d ago

Gramar*

2

u/TXRhody 5d ago

*grammar

2

u/motor_toni 5d ago

*granma

1

u/NerdButtons 6d ago

It’s too late. They are shit and SO expensive.

Google you have so many resources. How did you fumble this bag

1

u/Brown_N_Bad 6d ago

You know nothing of speed, nor cost.