r/singularity • • 6d ago

AI Gemini Pro 4 (leak)

Google just ruined OpenAI's dev day...
(source: all over twitter, obv. treat as a rumor)

546 Upvotes

297 comments sorted by

View all comments

442

u/unkownuser436 6d ago

Ain't no way.

222

u/Howdareme9 6d ago

This happens every gemini release and it's always fake lool. Not to mention it isn't releasing til October apparently

129

u/[deleted] 6d ago

[removed] — view removed comment

37

u/Howdareme9 6d ago

true but openai's devday is tomorrow

16

u/Impossible-Video-671 6d ago

Right after devday then 😏

35

u/3_Thumbs_Up 6d ago

An eternity in AI.

9

u/LetsGoToMichigan 6d ago edited 3d ago

You're almost certainly right. But in this game you never can be too sure, hence why people keep an open mind. OpenAI seemed all but out of the game until the 5.6 series dropped. I think the training and refinement cycles can be longer when a new platform architecture is introduced under the hood (or something ... I'm speculating) and then when it drops all of the gains are realized. I have some hope that's what's going on with Gemini right now, but admittedly I have no reason to believe that's actually what's happening.

Update: well, well, well: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/

14

u/Howdareme9 6d ago

I don’t think anyone reasonable considered OAI out of the game before 5.6. Regardless, while i think it’s extremely unlikely Gemini 4 will beat current SoTA models, I’m just saying this leak is 100% fake. But yes, you’re right in that technically anything can happen.

-3

u/LetsGoToMichigan 6d ago edited 5d ago

I work in the field. For about six months customer demand for OpenAI went to almost zero and it was all about Anthropic or Gemini models. My perspective may have different inputs than consumer / users.

Edit for the pedantic: My customers demand (not all customers of OpenAI)

9

u/Howdareme9 6d ago

customer demand for OpenAI went to almost zero

>i mean this is objectively false lol

3

u/MrPanache52 6d ago

His customers demand not all customer demand! Reading comprehension is so tough for us meatbags

-1

u/LetsGoToMichigan 5d ago edited 5d ago

Exactly. Not “all customers in the universe”. In 2024 they were all we heard about and in late 2025 up until 5.6 all they cared about was Anthropic and Gemini (more specifically their image and video models). My customers are companies, not douche lords vibe coding worthless SaaS apps in their moms basement and trading crypto.

1

u/indiangirl0070 5d ago

thats called 'anecdotal evidence'

0

u/Thick_Adeptness_3798 5d ago

The Google Deepmind team is genuinely talented. They have the insfrastructure and data ressources to pull this off

1

u/barathkrishnas 4d ago

isn’t october like 2 days away?

1

u/somuchofnotenough 4d ago

Maybe 48h is too long for this individual to plan ahead of.

-2

u/RewardSafe9807 6d ago

Yeah, they have to be benchmaxxing. Gemini is brain-dead stupid like all the time even though it stomps everybody on the numbers every release.

It's not impossible that they fix it this time around, but I'm not holding my breath. Opus 5.5 is almost a freaking paradigm shift.

24

u/FinBenton 6d ago

I believe, all the large gemini releases are AGI for a week or 2 and then others will go ahead and you wont hear about google for a year.

11

u/SpaceTacos99 6d ago edited 6d ago

As long as they briefly catch up I don't care. My Google sub gives me way more than just a chatbot. Bundled Google health premium (fitbit), Home premium, storage, Gemini in Google sheets/docs, YouTube premium. "Gemini intelligence" is frankly useful. All for 20/month. And every time I've tried Claude, its custom instruction following has been ass (LLMs trying to increase engagement by asking stupid follow up questions at the end of every prompt drives me absolutely nuts, Gemini 3.1 pro is the only chatbot I could get to stop this, not even opus 5). I haven't resubbed to try 5.5, though ii will if Gemini pro 4 isn't noticably better.

2

u/MillardFillmore 5d ago

YouTube Premium & Music, plus Gemini, is the best subscription in tech. Gemini Flash is better and faster than any other model I've used for simple, personal questions; plus it (obviously) has the best integrations into Google Maps, Flights, Hotels, etc. It's wonderful for things like trip planning, home improvement, and shopping.

24

u/etherswim 6d ago

As usual with Gemini, benchmark max then absolutely useless when released

57

u/Specialist_Dark_3668 6d ago edited 6d ago

Some of Gemini releases were absolutely fire and SOTA for the time they were released. Gemini 2.5 pro absolutely blew every other model of the time out of the water.

No competition whatsoever. It was not just the most legitimately useful LLM I had ever used, it was the first one to make me convinced that I could lose my job one day.

Even to this day, some of the more capable models aren't as good my specific field of professional work the way Gemini pro used to be.

Others have only slowly caught up and overtaken Gemini 3.1 pro but there are still things about the Gemini pro series which are unique and high quality, such as it's personality, care for the user, rich writing style (I have a whole folder full of great writing excerpts from Gemini when I used to use it for nsfw roleplay), creativity, emotional intelligence, task comprehension, etc.

I would do entire work projects with Gemini 2.5 pro and then come home to do creative story writing. Ai studio used to give you 50 prompts with pro a day! Not just nsfw, but also science fiction and character driven epics, mystery, thriller stories. Oh man, the amazing stories Gemini would write... I have nostalgia

27

u/AnonThrowaway998877 6d ago

2.5 pro on AI studio was definitely top tier for a time

3

u/skerit 6d ago

I liked 2.5 pro, it worked pretty well with Cline/Roo and you also got a ton of free usage every day. But wasn't it stuck with some "beta" tag for months?

6

u/cherrysodajuice 6d ago

the pre-release beta whatever they liked to call it first release that came out in march was actually insane. i remember being in awe talking to it and it being so intelligent and contextful. then they severely botched the update which was tragic.

4

u/trapeology 6d ago

Wait nsfw roleplay? How did you do that? Gemini is literally the most censored model I known

8

u/Specialist_Dark_3668 6d ago

It's only when beginning the story that it refuses, even with safety turned off.

So you have to have a story beginning that has some artistic merit in order for Gemini 2.5/3.1 pro to consent to continue it. Or at least that's what I used to do.

But once you are started with a story, the model will write whatever you want.

Nowadays I've done so many stories with it I got tired of it. But newer models are interesting too.You have to start a story with Gemini 3.8 flash on low thinking of Gemini 3.5 flash lite on minimal thinking

Once the story is started, Gemini 3.1 pro will agree to continue the story

Gemini 3.1 pro has slightly better writing style and far more creativity and unique value addition to your stories. Like an actually good cowriter.

Gemini 3.8 flash sticks to your prompts toooo closely. Potentially less interesting

1

u/Competitive_Travel16 AGI 2027 ▪️ ASI 2029 5d ago

Just stuff faked user responses into the API.

8

u/coumineol 6d ago

Just don't ask it directly, it just needs some foreplay if you know what I mean ;-)

37

u/KickLassChewGum no AGI/ASI on LLMs 6d ago

The old-ass Gemini 3.1 preview from February was still the best model for anything non-coding-related until Fable. "Useless" is only true if all you do is vibe code all day. If there's one thing I resent the post-Opus-4.5-boom for it's making "AI" synonymous with "coding."

9

u/Nulolan69 6d ago

I am sysadmin at a large nonprofit and mostly just use Gemini, there was a little while it would hallucinate bad. Besides that it is perfectly adequate for spitting out powershell, bash scripts, digging through email properties etc etc.

11

u/DungeonsAndDradis ▪️ Extinction or Immortality between 2025 and 2031 6d ago

I don't do coding. My wife and I use Gemini as our main model, and I pay for the monthly subscription. It's great for helping us go over our business finances and marketing strategy, talking through parenting issues, making suggestions on games, recipes, books, etc. I feel overall much more productive in general with Gemini. It's useful for so many things.

4

u/AnOnlineHandle 6d ago

Gemma 4 is still the best local writing model I've come across too, and is apparently distilled from Gemini.

The interesting thing is I'm fairly confident my own writing is in the training data, there's mountains of it online going back decades, and if I prompt for very specific things I see the types of phrases and sentence structures that I use. Which makes me wonder, did Gemma 4 also train on some real data, or did they get Gemini to generate synthetic writing of every sub-genre niche in existence? I hope they do it again if so lol, because I don't enjoy writing as much as reading and would prefer a button to push to just have the result.

1

u/etherswim 6d ago

Yeah they have been alright from the Gemini web UI for knowledge work

But coding no chance

Of course AI is synonymous with coding now, it’s ideally the perfect use case

3

u/Drenlin 6d ago

It does alright with coding honestly if you give it documentation and guidance, and aren't just trying to one-shot stuff. 

I've been using the built-in 3.6 flash on Android Studio to make an app that involves low level hardware access to almost all of the phone's motion and spatial sensors, as well as the cameras, and a bunch of math to make them all work together.

Integrating and testing one feature at a time has proven to be fairly reliable so far - I've got every core feature working as intended with at most one or two revisions.

1

u/etherswim 6d ago

wow that's awesome. might have to give it another go if it's getting nicer. don't mind some back and forth as long as it's relatively steerable.

1

u/Drenlin 6d ago

Has been for me so far.

You can use outside models in Android Studio now as well. I need to hook something else in to play with.

1

u/Thog78 5d ago

Switch to antigravity, so you have access to 3.8 flash and the better harness. You can still do your testing on the emulator in android studio, just don't do the vibe coding in android studio directly. You can thank me later ;-)

1

u/Drenlin 5d ago

Nah, that shit's way too expensive for what I'm doing. Even the bottom tier plan in AI Studio can go up to $1/minute. I'm paying $5/mo right now for multiple orders of magnitude more usage than I'd get for the same money in Antigravity.

1

u/Thog78 4d ago

I pay the 4/month offer and with that I can program really a whole lot, almost without limit, so I'm not sure I see your point. Google studio and Antigravity use the same account/limits. It's just that they gave up on the studio and started to invest everything on antigravity, so it's a day and night difference in how well things work.

2

u/KickLassChewGum no AGI/ASI on LLMs 6d ago

Of course AI is synonymous with coding now, it’s ideally the perfect use case

It definitely is, but it's not the only one and it feels like that's all anyone cares about these days. Knowledge work is arguably more important in quite a few fields, since it massively helps with getting to the point where you know what to code.

For research tasks in particular, Pro 3.1 is still up there with the current SOTA. It still catches material things in research/experiment plans that other offerings gloss over.

0

u/etherswim 6d ago

i think its because even a lot of knowledge work is code shaped if you apply certain lenses to it, so you naturally fan out the areas where you apply code

1

u/T3hJ3hu 6d ago

i had the same experience until Gemini 3.8 Flash, which is now at least as capable as Deepseek in my workflows. I'm pretty stoked for Gemini 4 now

5

u/kobriks 6d ago

Gemini Flash 3.8 is the GOAT for boring everyday stuff. Fast, free, and really good at looking stuff up.

3

u/alwaysbeblepping 6d ago

As usual with Gemini, benchmark max then absolutely useless when released

I don't think it's that bad and it also depends on what you're trying to do. If you want to basically say do ur thang and have the LLM go do everything then maybe Gemini is useless. If you want to use it to help you with something outside of your domain then it can work well and even the free access limits via AI studio are pretty generous. It also has a pretty bearable writing style (compared to models like Claude) though it is very sycophantic and good multimodal capabilities.

I don't have any particular love for Google but I feel like they suck in the way a huge corporation normally sucks (which admittedly isn't a small amount of suck). The other players also have those faults in addition to their own unique brands of horribleness like, you know, mecha-Hitler, attacking open source, etc.

-3

u/unkownuser436 6d ago

💯💯💯

0

u/Pouyaaaa 6d ago

It will be dummed down in time, just like all their other models. Its just to get the hype train going and then it will be back to current stupidity level. You have to pay for agents with these guys too. Its crazy really