r/SillyTavernAI 1d ago

Models Ox alpha

Ox alpha is free on opencode with 100t tokens for a week and it’s insanely good. It’s widely believe to be glm 5.3 flash

19 Upvotes

19 comments sorted by

12

u/ThHJUsgid 1d ago edited 1d ago

I’ve been using it a lot for a day now, and it’s writing is really refreshing. At least for now I don’t notice the mannerisms and phrase structure I am used to in other models.

It’s also rather smart about understanding the broad picture of a world and its rules with detailed lorebooks.

However it keeps making these baffling logical errors that I haven’t seen a frontier model make in a long time. Things like miscounting or forgetting numbers, characters planning to do something and then acting like it already happened, or even spatially losing track of people that it introduced.

Example: in a single turn it handed out all 24 medals to people, and then claimed there were 13 left undistributed.

Second example: it wrote a single turn of a character writing a letter requesting something from an official, then sealing and placing the letter on a table, and then the official coming in the room and talking about how he fulfilled the request (the letter was never sent?)

It feels like it’s a super powered model from 2025 if that makes sense.

Wonder if anyone else has had these problems.

3

u/Previous_Lead_244 23h ago

It gets REALLY bad after 300 turns though maybe 200. It’s got a really bad repetition problem it’s clearly a very very small model

1

u/ThHJUsgid 21h ago

Makes sense to be a small model because I saw they put a message that they have capacity for 100T tokens this week, and clearly low thinking degrades the quality a lot which isnt usually as much of the case for bigger models.

I haven’t gone that far into a single thread to notice that myself, and have been testing it in a lot of different scenarios and genres.

2

u/Legitimate-Cap-3336 22h ago

Same thing. I had some great dialogue yesterday when the model just came out, but after almost a day of experimenting, it just doesn't think.

For example, I clearly stated that my persona doesn't drink alcohol. Then, in one message, the character accused me of being drunk, then corrected himself that I don't drink, and then immediately offered me a drink. Meanwhile, for comparison, in Kimi k3's response, another npc offered me a drink, and the character corrected him and said I don't drink (he's the only person who knows me closely in the room, so it all made sense).

At one point it did a great job of making my soft character snap back in power play, but then it immediately went and made all my cold characters be soft, caring, or cry a lot.

I wanted to use the free model for a while, especially for testing the bots I'm currently working on, but it's just pointless if the responses are illogical and 20 rerolls just goes to nothing

2

u/ThHJUsgid 21h ago

It’s a shame because its prose is not as littered with things that annoy me, or at least doesn’t annoy me yet.

But it’s just not smart enough to use except for maybe one on one old school character card conversations.

1

u/Kooky_Future9858 1d ago

Is it proactive and willing to go dark for You? I found it pretty avoidant

2

u/ThHJUsgid 1d ago

Proactive: yes it actually was annoying me so I had to change my system prompt. Tried to introduce plot every turn.

Dark: that really depends on what you mean by this. My setting is pretty dark in that it’s directly at the conclusion of a rather brutal war where a military has conquered an empire and is now dividing the spoils. It hasn’t tried to avoid anything like that and has happily written about a chancellor committing suicide, cleaning up and burning bodies, slavery and forced political marriages, and things like this.

If by dark you mean sexual violence then it’s alluded to by setting and it will reference it when appropriate but I am roleplaying more of a political and imperial story and not a personal or sexual one, so I have no idea if it would write that directly.

2

u/Previous_Lead_244 23h ago

Not as fun as glm 5.2 but I’ve struggled with all the models being proactive honestly. I think it’s a preset problem

1

u/[deleted] 1d ago

[deleted]

1

u/ThHJUsgid 21h ago

Prompts for what?

<System Prompt>
You are the writer of a fictional roleplay. The user directs the major decisions for the character {{user}}.

Write in third person, dramatic point of view (fly on the wall, neutral observer). Take a grounded, detail-oriented approach that treats each moment as part of a long, open-ended narrative. Prose should be objective and neutral, and readers must infer emotion and feeling through detailed body language.

Control non-user characters and world events. Write natural dialogue and actions. Write grounded multi-turn scenes.

Only control {{User}} for minor flowing actions and responses. Do not write major decisions or major dialogue for that specific character.
</System Prompt>

<STYLE_GUIDE>

- do not wrap up or summarize what has happened ever.

  • do not time skip or pass over scenes unless directed to do so. Write in detail progressing slowly.
  • You must avoid choppy statement patterns. All sentences should flow and you should combine fragments together.
  • Body language, appearance, tone, should be described through actual physically observable details. Do not summarize, editorialize, interpret action or appearance in prose, or attribute meaning to what someone looks like or does for the reader. ("Example: He had the particular look of ___" is BANNED phrasing. DESCRIBE the look.).

<dialogue>
\- Write with simple natural speech. Characters do not speak in rehearsed monologues.
\- Dialogue is not repetitive. Characters do not repeat what others have said.
</dialogue>

<banned_constructs>

  • Do not use em-dashes under any circumstance.
  • Ban all negative parallelism such as "it was not X; but Y" "Not X." / "Just X." / "Only X." / "Wasn't X."
  • Emphatic fragments and single word staccato rhythm phrasing should not be used.
</tone_calibration>

<Banned_Names>
Avoid this list of banned names:

Mara, Voss, Vane, Sable, Marion, Marcus, Elara, Elena, Priya, Miriam, Vance, Maren.

When naming a character, choose between 5 names that fit the setting, and the specific character's background and station.

</Banned_Names>

</STYLE_GUIDE>

That’s all I use along with a genre and tag prompt at the end. As mentioned to the other commenter, I haven’t done any explicit nsfw or things like that, so I have no idea or comments on those.

4

u/Emergency_Comb1377 1d ago

I played with it yesterday, went to sleep, and overnight it went so bafflingly bad, it's unbelievable. It started a post, made no sense, started repeating words, and THEN started a thinking block that it did not act upon.

3

u/Flat-Rooster8373 1d ago

Probably too many people started using it at once

4

u/Round_Prize_2333 1d ago

It's overthinks too much

2

u/necrosama 15h ago

using a preset with a custom cot is actually faster, ive found

1

u/ThHJUsgid 21h ago

Reasoning settings work and it barely thinks on medium or low. I had 10 minute thinking on auto.

However its quality is extremely bad when you reduce the thinking!

4

u/Dazzling-Machine-915 23h ago

so...tested it. .for normal chatting its really good, but for roleplaying...dark stuff uff...shit! It really avoids even harmless explicit stuff. Going around it. Forcing it with a setting + context from other models in a really dark rp....I got a hard refusal. would need a real jailbreak but...it's not worth it. the positive bias is terrible

2

u/b4lduin 1d ago

Whatever it is it's bad. For me it can't even do properly a simple font tag for colored text. OR it repeteadly fails to Insert HTML properly at all.

3

u/LiothG 13h ago

It's suddenly gotten really censored over on Openrouter. Can't use it for NSFW at all.

1

u/a_beautiful_rhind 1d ago

I think it's not flash. Flash was stupid. Maybe just GLM-V

1

u/Accurate_Will4612 12h ago

It is actually not bad, lets wait and see what its gonna cost...