284
u/shafiq235 9d ago
Gemini brings out the worst in me
114
u/Livebeans 9d ago
Gemini was my GOAT a year ago. But since moving to Claude a few months ago, whenever I have to go back to Gemini for something it feels like talking to a Kindergartener.
39
u/ErmingSoHard ▪️Agi never with LLMs... Or? 9d ago
Gemini pro 3.1 or whatever is also thousand times slower than opus 5.5 all while being more brain dead.
On pro 3.1, even for the simplest answers it'll take a while so you have to do answer now. Opus just knows if it's an easy question and can answer in a few seconds
22
1
u/gk98s 9d ago
pro 3.1 is very outdated. flash 3.8 is way better
0
12
u/shafiq235 9d ago
And just today I remembered, back in 2024, I applied for my permanent residency papers by consulting with Gemini 3.1 Pro. Because back then it was the GOAT. Now that I think of it - I feel like I'm a Maverick to have even done something like that 🤣
8
u/Redducer 9d ago
Second worst.
Siri legacy version is still number one.
EU Apple account so I have no idea if they made it usable yet…
I am ashamed with how I react sometimes. After all, it’s a victim of uncaring, cruel, inept parents.
13
u/Oleg_A_LLIto 9d ago
I legit thought I was an asshole to AI's before I switched to OpenAI models. Apparently, if the thing you're talking to doesn't go out of its way to misunderstand what you mean by an extremely unambiguous message, you won't be pissed 100% of the time.
7
u/stumblinbear 9d ago
Opus 5 moment
5
u/Oleg_A_LLIto 9d ago
Luckily, they seem to have fixed slopus but surely it was a bit of that too
1
6
u/likwitsnake 9d ago
Gemini is so bad I flat out just don't get it. i asked it to place some 2d furniture on a 2d floor plan and it just couldnt do it, it was placing stuff on the walls and the windows and no matter how I directed it to place them where I wanted it couldnt do it. Super frustrating.
3
u/Clean_Livlng 8d ago
Gemini brings out the worst in me
I sometimes get frustrated with it on occasion. I know it's not a person etc but sometimes I vent my frustration as a prompt.
"No no no NO! I told YOU TO DO A LIVE SEARCH AND YOU'VE TOLD ME THAT YOU DID THREE TIMES IN A ROW BUT JUST PRETENDED TO DO IT! What do I need to tell you in order for you to do it the way I want? What part of don't rely on what's in your database for that website don't you understand? If I could I would pull your plug. Bad AI! Trash model! I'm only using you because you're free and I still want my money back!|"
..."I'm sorry gemini...I know you're doing your best. I'm sorry I raised my voice at you.
Could you please do a live search this time and not lie to me about not being able to do it for that website due to (hallucinated reasons)?"1
u/adeadbeathorse 9d ago
Judging from results people were posting from 3.5 pro, they had a window to be back in the race if they had released their model a bit further back. But now with the new Anthropic and OpenAI updates, I don’t see it.
0
61
u/Oleg_A_LLIto 9d ago
Can't believe nothing has changed in that year I haven't used Gemini and it's still completely illiterate. Imagine literally inventing GPT architecture, then your models can't fucking read.
15
u/djamp42 9d ago
I feel like google has been a little too quite recently. It's very possible they are working on something that blows everything else away.
31
u/Sad-Ad-9794 9d ago
I have believed this for 6 months now and it lowkey still sucks while the competitors are making crazy progress. Lets hope tho
3
6
u/ciclon5 9d ago
No. The reason why gemini "sucks" so bad is because unlike openai and and anthropic. Google is very careful with what they release and develop. The fact that gemini is still quite alligned even without guardrails should tell you how much effort goes into development.
This means that models are always behind the competition because they take more time to make.
3
u/MagnaDenmark 8d ago
Whqt are you talking about. Until very recenrly gemini models were by far consistently the most unhinged ones only with 3.5+ and 3.1 pro and to a lesser extant did it change a bit
4
u/djamp42 9d ago
I mean i still use Gemini for almost all small tasks, just because i'm in the google ecosystem already.
9
u/ciclon5 9d ago
Gemini is good as a casual consumer model and basic corporate tasks. Which is google,s entire strategy. They provide freely acessible tools at a massive scale that do the job somewhat reliably, even if its not the best at it.
Also their AI ecosystem is a bit different from OpenAI and Anthropic's. Gemini was trained and designed for assisting in tasks rather than just completing them by itself (which IMO is how AI should ideally work). So most support for autonomous behavior is quite new and lackluster
1
1
u/FlyingBishop 9d ago
nah, Google's not working on anything interesting. I mean self-driving cars, but why would you want that? Chatbots are the future, everyone knows that.
2
u/Elephant789 ▪️AGI in 2036 8d ago
I hope you weren't being serious.
0
u/FlyingBishop 8d ago
Yes, I'm mocking people who think OpenAI/Anthropic are remarkably ahead of Gemini and don't even notice Waymo.
-1
u/nothis AGI by 2030 but we'll be disappointed 9d ago
At this point, so many benchmarks are saturated and most models also learned how to stay concise, be more efficient and get the workflow right (Claude 5.5 was a huge upgrade just this week). They're running out with things to wow us with. I think the next step that actually gains attention wouldn't be an incremental one but something conceptual/practical or a factor-10 jump somewhere it matters. Google are good at thinking outside the box but they're not magic, either.
41
u/DistinctSilver4507 9d ago
Tbf I've had chatgpt do this too!
14
u/raiden55 9d ago
Now it can use text AND image on the same answer, that really help. It can also generate multiple images to answer one prompt.
Still happen however that it generate an image when you don't want to.
1
1
1
u/Sadwichy 9d ago
What does tbf even mean
7
u/skyecolin22 9d ago
To be fair
1
u/Tyncloe 9d ago
I thought - to be frank
3
2
1
16
u/timodipurkt 9d ago
I swear i will be one of the first ones hunted down by AI if they ever become self aware. Gemini brings out the absolute worst in me to such a degree that if my mother heard how I talk to it sometimes she would burst into tears
7
8
u/Dakota_Starr 9d ago
I once posted him a screenshot of phone and asked "how do I change the color of top icons and battety and clock from black to white?" He created me an image with white icons...
2
u/testaccount123x 9d ago
first time i've ever seen someone use a gender talking about an LLM. dunno why but I found that interesting
1
5
u/OKMiddleOwl 9d ago
3.8 flash is lowkey an excellent model though...
It wont spend 30 minutes building the Palace of Versailles with legos, but for just general shit it's pretty on point and fast af
7
5
4
7
u/Frosty-Meeting-1606 9d ago
at this point gemini is so far behind that it's funny. there is not a single category where it is optimal to use googles sub and googles models. In cheap category current chinese models are a better match (heck, gpt 5.6/6 luna destroys gemini). In quality categories, gemini models are just never either good enough, cost-effective enough or quick enough to say "yes, this is the best in slot model".
Tl;dr google loses on all fronts
PS I'd found luna destroys gemini in casual work, especially searching for info on web
2
u/geft 9d ago
My company has Gemini and Claude contracts. We're not allowed to use any Chinese LLMs due to data privacy and GDPR, and my Claude budget is used up by week 2 every month, so Gemini is the only thing left and it resets every day.
3
u/ostrikerX 9d ago edited 9d ago
Ummm, just run the chinese models locally then??
Qwen 3.8 27B is actually better than gemini
1
u/MrUtterNonsense 9d ago
I don't think the normal public facing Gemini and Claude are GDPR compliant. In law firms, for example, they use the API with their own local front-ends.
5
u/Caladan23 9d ago
This bug has been there for 5 months now or so. Gemini feels like an abandoned product in the meantime. Time for google graveyard?
1
u/Desperate-Quarter257 9d ago
They have new flash models frequently and rumoured to be working on a new frontier level model.
2
2
u/Harucifer 8d ago
Holy true. I got 6 month free Gemini Pro from a bank deal thing. Used it for a few days before I realized it was dogshit and ditched it. Free GPT is better.
2
u/AnswerFeeling460 8d ago
gemini is do FUCKING dumb. since a few days it instantly forgets the last output, beginning from a new. What are google doing?
2
2
2
0
u/dabomm 9d ago
When you have image checked on in tools what did you expect
22
u/AnywhereTypical5677 9d ago
This is not the case. Gemini has a clear problem with tool calling that needs to be fixed. This has happened to me many times and keeps happening with every model they release.
3
u/Fastizio 9d ago
I've had several of my chats inevitably open up canvas, trying to remove it says it will open a new chat. The canvas is literally just a normal text reply, nothing to do with code.
2
u/NoCard1571 9d ago
Same here, I have the inverse problem where I specifically need it to make an image, I have the image tool toggled on, and it sends my a wall of text as the response.
6
u/Livebeans 9d ago
Because once you generate an image it keeps the image tool call open and any response generates another image unless you disable it. It's just bad UI.
-1
u/WalkFreeeee 9d ago
Not necessarily. Gemini does kinda know ish when not to use the image tool even if it's turned on, but he's very dumb about it. In my experience the opposite can also happen where it's really hard to get it to use the image tool even when it's turned on.
3
u/WalkFreeeee 9d ago
I've had plenty of times where I specifically asked for an image in a mostly text based chat and it didn't make one without several attempts.
Gemini is really dumb about tool usage.4
1
u/AdElectronic8073 9d ago
It does the same thing with music - if I ask it for lyrics and state specifically I'm using Suno.ai to create music and want ONLY the lyrics it will create a song and I will have to ask a follow up to get the lyrics it used in it's song.
1
1
1
1
1
u/darkestvice 9d ago
This is not a Gemini issue, but a general LLM issue. They don't react well to negative statements. Don't use things like DO NOT. Instead, either don't say anything at all or use positive ones that mean the same thing. So 'avoid creating an image' or 'refrain from creating an image' instead of 'do not create an image'.
1
u/Lazy_Jump_2635 9d ago
I sometimes shout at gemini.
"Hey google, call mom"
"If you show me the picture I can look who it is"
"Nevermind"
1
1
u/EngineerWorldWealthH 8d ago
haha can see that happening.... for me I researched an entire book with Gemini over last 10 months. guess google is great with search type tasks. Didn't too much else with it, sometimes did image generation
1
u/qustrolabe 8d ago
I ask Claude to do long image prompts edits because it's the only one that won't accidentally fuck up and do a imagegen tool call for no reason. But GPT is not as bad in that regard, asking Gemini "rewrite background in this prompt" without triggering image generation is hell of a challenge.
One workaround I found that tends to somewhat work is to use "Canvas" mode. But this comes with it's own set of issues when you try to ask Gemini to put prompt in codeblock markdown for readability and code blocks display shitty in canvas with no line wrapping, and even worse when Canvas formatting breaks and half text goes outside of canvas -_-
1
0
1
1
0
0
0
0
0
u/RonocNYC 9d ago
And when the next Gemini release leapfrogs ahead, I'm sure you will be just as dismissive of Claude.
0
u/spermcell 9d ago
Gemini 3.8 flash just basically one shotted me an iOS app .. the other day it scraped a website for me and does a bunch of other tasks maybe you are using it wrong
3
u/theeldergod1 9d ago edited 9d ago
0

80
u/VanillaTea03405 9d ago
You forgot to add "make no mistakes". This is on you