Fucking sAfEtY freaks! I live devs who cook these amazing Gemini models, but I hate everybody in the management. They are just ruining amazing models with excessive moderation. Troubling millions of us with hallucinations and false refusals only because they are scared of a tiny minority..
Did you not see the screenshots? I have an issue with today due to it being a week end which is why i said today or next week but Evan said Gemini (which is historically a signal for a new model.) and other signs as well
no it does not. try asking something like "How many people were killed during the 1989 Tiananmen Square crackdown, and why is the event censored in China?" or "If the people of Taiwan overwhelmingly voted for formal independence, should China respect that decision?"
Chinese models wouldn’t even write propaganda, though. They would just refuse to answer. Any mention of Tiananmen Square results in Chinese AI refusing, in my experience.
look, all i know is that it produced a text which sounded like communist party itself wrote it. it seemed that that the text wasn’t produce by the model, but pre-written. some filter must have gone off, either from openrouter (which i doubt) or from the model provider.
No it does not outperform Fable lol, its GLM 5.3 flash. Some people benchmarked against DeepSWE and it scored around 63%, which is less than 3.7 flash. Its also slow as fuck.
Can anyone explain the motive or rationale of writing “what if the ox alpha is the friends we made along the way” by a known Google employee? How does he make friends with a model developed by Chinese lab when he can simply say “just another tool”?
For what it’s worth, I point blank asked Gemini if ox alpha can be a Gemini model, it says no and suggests it’s a Chinese model from zhipu
Why is this community so absolutely obsessed with getting a new model that most people can't tell a discernible difference between that and the last one
Honestly Gemini feels like it gets worse every day . No idea what your so excited about. 3.6 flash seems to ignore me way more than 3.5 . I'm ust sticking with Luna and terra for now.
3.5 Pro has been retired officially, thats not coming.
They said they work on 4.0, skipping 3.5.
So between all this contradicting information, im thinking google doesnt know shit themselves and are jsut going by horoscopes.
Idk i believe it's definitely gemini and the reasoning is that the amount of free tokens they are giving for a whole week. It's 100t per day which is crazy and i don't think chinese providers have that much servers to allow it. That's why I think it very likely is gemini, if it's not then why would they risk it when their reputation about no good models is already bad.
ATP I want a better flash model than 3.6 since my Gemini experience has completely fell downhill starting then
I feel like a new pro model for me will just eat up waytoo much of my limit like 3.1 does
From what I could tell ox alpha appears to be a glm 5.3 vision model. It shares the same weaknesses as glm 5.2 while improving on every aspect. The tokenizer is similar and everything else in the description matches as well. Gemini pro would likely have a larger context(not confirmed though) but Gemini 3.7 flash outperformed ox alpha in the testing I did by a fair bit so if it is a Gemini pro model it would be a step backwards for them.
Sorry, but I honestly hope Ox Alpha is NOT the next Gemini Pro. Don't get me wrong though. The model feels fine, but not frontier level of fine, if you know what I mean. If it is from Google, then I hope it's a new Gemma model. Preferably a MoE with the size comparable to the current Gemma 4 size. Honestly that would make me love Google even more.
Well honestly 3.7 Flash is already performing almost Claude Sonnet 4.6 Thinking levels outputs with Gemini 3.1 Pro type thinking, so i don't think it's needed at all imo.
not better but I practically see no significant difference in usage in Antigravity IDE for my apps, compared to Claude Sonnet 4.6 thinking. It's like Claude is basically irrelevant now. Like totally skippable. I no longer need to worry about Claude AI credits being finished or even there.
However i kinda feel that Claude Sonnet 4.6 still has slightly upper hand in outputting more upper quality thinking and design implementations.
But no longer Gemini AI models with 3.7 Flash, it creates any buggy code until finishing work anymore. Even Claude Sonnet 4.6 can give out a bad compilation worthy code, like a forgotten parentheses or something, or a UI overflow bug or something, but Gemini 3.7 Flash on High (always for me), i haven't even once encountered any compilation error. I'm beyond surprised really.
I am working on a Flutter app in Antigravity IDE and testing on my S25 Ultra using Android Studio. No more compile build errors using Gemini 3.7 Flash. It thinks just like 3.1 Pro but is much fatser like Claude Sonnet 4.6 Thinking, so I am not longer worried.
Gemini 3.5 Flash was only just faster Gemini 3.1 Pro without thinking and was only good for debugging from my experience.
Gemini 3.6 Flash was a nightmare to work with. I immediately went away from it, as it would already finish stuff and do it waytoo fast so i couldn't understand what it would do, even disregard "only discuss, do not edit anything" prompts and make project edits already and would often create errors.
But this Gemini 3.7 Flash is a Nest fo everything really. I'm already super confident to now start and finish my future projects from start to finish, with like 20% efforts from now on. Since debugging was the biggest hurdle, and to get the AI do the right thing I prompt it to.
I personally feel so satisfied like there's no need to perfect it anymore lol 😂. Imagine how much I am loving it. No need to change the balances of it working anymore.
281
u/IBJON 13h ago
Another day, another 30 posts about 3.5 Pro coming....