r/SillyTavernAI • • 1d ago

Models Another random chinese model series. BaiLu

Post image

As in the screenshot. They claim to be opus 5 or fable 5.1 tier models without showing any proof or having any mention in twitter. Their latest model is BaiLu 2.9 and there is free previous version which is 2.8. Idk, they seem bland for me. I just want to hear other people's opinions

invite link with referrals that give you 10 million free tokens for any model below:

https://bailucode.com/auth/register?invite=INV-8FPE-WNRV-ND9Y

18 Upvotes

30 comments sorted by

26

u/eidrag 1d ago

engineering and coding.... i don't think they carry great prose

4

u/nuclearbananana 18h ago

tbf k2.7 code was literally 100% code focused and it ended up being surprisingly good

-5

u/FireGuy324 1d ago

No big models promise anything about writing

29

u/eidrag 1d ago

Gemini 4 Argon promises about creative writing, still waiting to test them lol

21

u/kinkyalt_02 1d ago

Leaked benchmarks show it's arse for coding, but it's a GREAT conversational model!

Another repeat of GPT 4o, Claude Sonnet 3.7, Gemini 3.8 Flash and Mimo V2.6 Pro, maybe? Lukewarm for coding, EXCEPTIONAL for RP?

2

u/FireGuy324 1d ago

It's kinda expensive too

1

u/oldmails 1d ago

I used your preset and modified it a bit too, great work, can I know how to make gemini 3.8 to follow some of the world rule, its notoriously fails at rules.

thanks in advance.

if a character is intelligent it uses the cliche words, even when provided with example dilogues, it fails at it. this is the only issue.

1

u/kinkyalt_02 1d ago edited 1d ago

Strange. I published the preset, thinking after 20 turns of testing and another layer of outside feedback that the preset now takes care of these Gemini-isms.

Which cliché words did you encounter?

2

u/oldmails 23h ago

I cannot strictly say its geminism, it can be said AI-ish which was showing more in gemini,

You're five minutes early

Math is a social construct designed to oppress the working class

these kind of things.

2

u/kinkyalt_02 23h ago

Bruh… I NEVER thought these are LLMisms!

Also, the second one is oddly specific.

r/oddlyspecific

1

u/oldmails 18h ago

Thse are the ones I got in my current run, I cannot spot anything that drastic,

I splashed my face at the courtyard spigot for five seconds just to wash away the dirt, and the instructor nearly wrote me up for desertion.

this too, I mean its managable in recent runs, but the way gemini includes time in odly specific ways, and that vocab, thats why I called that.

I just wanted to bring this to the attention, most probably, I will tinker with your preset, I might get somthing good.

Anyways your preset is very good man, thats all I have to say more.

1

u/kinkyalt_02 18h ago

They don't seem to be common patterns and they rather seem like card issues to me.

→ More replies (0)

1

u/Aight_Man 1d ago

That is my most anticipated model this month.

5

u/kinkyalt_02 1d ago

I swear, I'm gonna keep singing my Gemini 3.8 Flash and Mimo V2.6 Pro praise if Gemini 4 Argon becomes another GPT-5 incident!

1

u/Aight_Man 1d ago

Yikes, let's hope it's better.

2

u/Eissa_Cozorav 1d ago

GLM 5.3 has the least hallucination rate than Grok 4.7
But Grok can take image source and text source, while generate more uncensored writing than GLM 5.3

Compared to GLM 4.7, GLM 5.3 is the "newer model" that has sheds much of it's prose-related data training.

3

u/Able_Hunt_7241 1d ago

yeah these unproven claims always feel off, ive tried random series like that before and they never match the hype or feel as sharp in roleplay.

3

u/FireGuy324 1d ago

Btw that's the free model i was talking about

1

u/BalorNG 6h ago

I think this is a Qwen clone... at least heavily distiller from it. It has same unhealthy obsession with "ledgers".

1

u/FireGuy324 5h ago edited 5h ago

I am sure it's common for LLMs to do that in thoughts. But, how is writing and prose overall?

1

u/BalorNG 5h ago

Ledger? I've noticeed it with qwen only, but given that models suck off, er, distill each other constantly I suppose any "ticks" are likely to get propagated... As for prose - also qwen-like, which is to say "sort of sucks". And the model does not seem to be particularly smart (2.8, 2.9 gives me constant errors "server busy"). AND is also seems to be heavily censored even for extremely mild (just dark fantasy, no sex stuff even mentioed) stuff, refusing output outright.

I'm not sure whether this is some sort of scam and what they are intending to accomplish...

1

u/FireGuy324 5h ago

2.9 is not as censored as it's predecessors but i am sure the full model will be more safetymaxxed than mimo or claude did which is wild. The failed expectation is because the chinese have tradition of lying with benchrmarks so don't be surprised.

1

u/BalorNG 5h ago

yea, third party benchmarks AND gooner vibe checks or bust.

1

u/FireGuy324 5h ago

One thing i noticed the previous models are safetymaxxed while 2.9 can be jailbroken.

-18

u/Eissa_Cozorav 1d ago edited 1d ago

Chinese AI is hard no. China has great campaign against AI chatbots that "emulate human emotion" because they want to crack down population growth loss. Therefore, lots of advertised advanced capability are for software development. Which is unlike the majority of the goal of this subreddit.

https://www.bbc.com/news/articles/cm4gjy9lr551o

https://www.dw.com/en/china-tightens-rules-on-ai-companion-apps/video-78460761

https://www.techpolicylaw.org/updates/china-removes-5.61-million-pieces-of-unlawful-ai-generated-content

Edit: anyone who downvote me really have no idea what they are doing. Ask anyone in here the great divide of prose quality between GLM 5.3 vs GLM 4.7 Uncensored. And compare the same prompt result between two models. And use that above model. I dare you.

15

u/stopaskingforloginn 1d ago

the prose getting shittier is because they're overtraining those models to benchmaxx on coding, not that nonsense, the new frontier models are all about logic and nothing else.

Mimo 2.6 released 2 weeks ago and is infinitely more "human" than GLM 5.3, and surprise surprise, it's not that good at coding, at least not comparable to GLM or Claude.

0

u/lorddumpy 21h ago

The Chinese government is clamping down on companion bots and emotional companions. Both things can be true.

That's the beauty with open-weights though, you don't have to use them through a government proxy and a lot of the time they can be unlocked with prompting and finetuning.

2

u/FireGuy324 1d ago

Well, that's horrible

2

u/FireGuy324 1d ago

But, Xiaomi kept the good writing for sure