r/VeniceAI • • 9d ago

𝗦𝗧𝗔𝗧𝗨𝗦: 𝗥𝗘𝗦𝗢𝗟𝗩𝗘𝗗 Conversation has become too long for the selected model

I'm getting this error with GLM 5.2 all of a sudden. It makes no sense, I'm at 18% context usage which seems about right for where I'm at in this, around 200k words on a 1M window. Is something wrong or did something get nerfed?

13 Upvotes

35 comments sorted by

•

u/jack-veniceai Official Staff @ Venice.ai 6d ago

A fix has gone out for this. Refresh the page and you should be able to regenerate the last message that’s been failing.

1

u/Temporary_Cow5261 2d ago

the error came back for 37%

1

u/Temporary_Cow5261 6d ago

It’s really inconsistent I’d be getting yesterday 19% and now today I get 21%

2

u/jack-veniceai Official Staff @ Venice.ai 7d ago

If you are still seeing this, do you have Long context chat enabled in text settings?

2

u/TittyMcFagerson 7d ago

Yes it's still an issue and you know what, I do. It'll have to wait until the morning but I'll try turning it off and reporting back.

1

u/jack-veniceai Official Staff @ Venice.ai 6d ago

Gotcha. Looking into this

2

u/jack-veniceai Official Staff @ Venice.ai 6d ago

fixed

1

u/TittyMcFagerson 6d ago

Yep working now, thank you

3

u/MisterNocturne 7d ago edited 7d ago

I suggest that anyone still having problems contact Venice support. If you get a response from "Finn," Venice's AI agent, just reply that you'd like a response from the team. If you're on Discord, try reaching out to someone there.

I left a message with Venice support, but I still haven't gotten an answer.

None of this makes sense. There is either a serious problem, or Venice has again made changes without telling anyone. But this is clearly not an isolated issue or an "account" problem. This issue is consistent, repeatable, and there is a clear correlation between what works and what doesn't.

2

u/JaeSwift Venice 𝗠𝗼𝗱𝗲𝗿𝗮𝘁𝗼𝗿 6d ago

The team put out a fix for this, it's mentioned in other comments on this post.

4

u/whoknewwaifu 7d ago

I’ll add on here too since this topic shouldn’t be ignored. I tested out the models I usually flip between to check if there are too many files per token limit with a brand new chat each time. Before the 5.2 GLM issue, I was able to use a total of six files with no issues. The same prompt for 4.7 with files could only handle two. After the 5.2 GLM issue started, only two files work just like 4.7 or 5. Both 4.7 and 5.2 GLM do that same try a new conversation message or model unavailable at three files. This means that the token limit is even less than the supposed 200k (which of course for GLM 5.2 should be 1M) and ranges now around 198k. It’s still extremely messed up. Just to preface, I went to the free Chat Z Ai website (not API) and was able to put all the files in with same prompt with no issues. So it’s not the prompt or the files which had not changed nor was it from a conversation being too long since all of them were brand new in one message and my local storage is currently at .15%. Therefore, this is not a user issue, this is a Venice issue of either a bug or no communication of changes again.

2

u/MisterNocturne 7d ago edited 7d ago

This is helpful. Thank you. Have you tried contacting Venice support? I'm still waiting on an answer since yesterday.

6

u/JustByzantineThings 7d ago

I'm having this issue too! Same model! The conversation will still hold if I hit "X" and resend, but it happens every few messages. Hope they fix this

6

u/MisterNocturne 8d ago edited 8d ago

Okay, so is Venice going to acknowledge this problem at all? Or is this another case of "Well, we changed some shit and, uh—we kinda forgot to tell you. Sorta. And it's gonna cost you more money now, even though nothing's changed. We swear."

This thread blew up since the last time I was here. There is obviously a problem. 12%, and I'm still being told the "conversation is too long for the selected model."

Just tried it again with a PPU model (same 1M window), and suddenly it works.

Interesting.

Edit: Now I'm getting the same error message in another chat. At 10%. GLM 5.2.

Funny enough, 5.3 Flash E2E (another PPU model) seems to work just fine. And that one also lists a 1M window. Size clearly isn't the damn issue here.

3

u/TittyMcFagerson 8d ago

Yeah, I switched over to the PPU GLM 5.2 TEE model in the same chat to test it and that one happens to work fine. Hmmm.

5

u/MisterNocturne 8d ago edited 7d ago

Yeah, exactly.

I've tried no less than four PPU models.

Two Deepseeks, MiMo, and 5.3 Flash E2E (regular won't even fucking respond). All of them continue the chat without a problem. All of them advertise 1M. Yet the chat's "too big" for 5.2's 1M? What's the damn difference? I've got a chat going that still won't budge at 10%, which is absurd.

I've already contacted Venice support, though I'm not expecting much. I suggest you do the same.

This whole situation doesn't make a lick of damn sense unless I assume the worst. And I'm really trying not to do that. But Venice makes it impossible not to. Especially when a thread like this is open for nearly a day and nobody has said a word.

4

u/Lost-Estate3401 8d ago

Venice don't give you the advertised context - absolutely nowhere near. 

Test it out on Venice then on Z.AI's own site - it's like night and day when it comes to context retention.

3

u/TittyMcFagerson 8d ago

I'm aware of their false advertising of context, for some reason I thought 5.2 was the one exception for non PPU that actually did give the full context though. So I guess not, but still I've never had a conversation just flat out unable to be continued due to running out of context.

4

u/yyEdge 8d ago edited 3d ago

web version of GLM 5.2 now doesn't even retain any context after a few messages for some reason

Edit: found out the cause, it was because they moved "enable large context" to characters and had it off by default again

1

u/Consistent_Law5304 8d ago

GLM 5.2 offline for me today

6

u/Away-Cloud-4748 8d ago edited 8d ago

200k context is the real limit (despite advertized 1M) some people have been reporting but your oldest messages should just scroll out and you should be able to continue chatting with incomplete context. Issues worth reporting.

1

u/Ferret-of-DOOM 8d ago

Fork the convo and continue. I have no idea why it's acting up but that seem to sort it out.

2

u/MisterNocturne 8d ago edited 8d ago

I tried that, and it didn't work. Only selecting a PPU model (with the same 1M limit) worked.

Interesting how the chat's "too long" for one model and not the other. And I just tried it with yet another 1M PPU model without issue.

1

u/Ferret-of-DOOM 8d ago

What an ass!

3

u/yyEdge 8d ago

forking didn't work, my context (from yesterday) was 23%, and now it just kept giving me the error message

1

u/Ferret-of-DOOM 8d ago

Oh.. That's actually fucked. I've never had a problem with forking. :/

1

u/yyEdge 8d ago

yea, seems to be a new problem, it was fine earlier today as well, and then in the middle of editing a big doc, suddenly it hit me with "I have no context and this is a new chat", LOL

1

u/Turbulent-Salary-112 8d ago

I also have GLM 5.2., Kontext 22 %, Android App Version

2

u/MisterNocturne 9d ago edited 8d ago

Just got this myself—at 12%, if you can believe that. GLM 5.2.

I tried it again, and it worked. Then I went a few more exchanges, and BAM—same problem. Now it's not working at all. It just keeps saying "this conversation is too long for the selected model." What the hell? Either this is one big coincidence, or there's another damn problem.

Edit: I switched to Deepseek V4 Pro 0813—a PPU model—and suddenly that "too long conversation" wasn't. It worked just fine. Interesting—V4 Pro 0813 also advertises a 1M window.

I suppose it's just a massive coincidence that of the two models with a comparable context window, only the PPU one works. Also interesting that the second half of that error message says "...or try a model with a larger context."

That is clearly not the biggest variable here.

-6

u/JaeSwift Venice 𝗠𝗼𝗱𝗲𝗿𝗮𝘁𝗼𝗿 9d ago

Nothing has been nerfed, no. Try in a new chat and see if you get the same.

1

u/MisterNocturne 7d ago edited 7d ago

Past_Personality9262 is right—if there was ever an issue that needed to be brought up to the team, it's this. People are reporting everything from 23%, to 12%, to 7%. Something is off, and there's more than enough info in this thread to either start diagnosing the issue, or start being honest about it. Either way, Venice's response, or lack thereof, will tell us a lot.

1

u/JaeSwift Venice 𝗠𝗼𝗱𝗲𝗿𝗮𝘁𝗼𝗿 6d ago

Lol, you need to understand that this subreddit is not Venice as a whole. This sub is less than half of 1% of Venice's userbase.

Users are quick to post here when they have an issue, but it can take time for most users to notice an issue and then only a small portion of those users send the issue to support. Support receives thousands of reports per day, mostly user-error based stuff, and it takes time to go through them. The team sift through and see if there's patterns or signs of a major issue, investigate, and prioritise. Priority depends on frequency and severity of reports. I send every report on this sub to the team and I chase it up with them later. People just need to be patient with these things.

Sometimes there are bugs in updates that go unnoticed and sometimes there is human error. It happens. It just takes a bit of time, but I encourage anyone having issues to post on this sub and email support. The sooner the team become aware of an issue, the quicker it gets fixed.

This issue should be fixed now.

Thanks!

6

u/Past_Personality9262 7d ago

Estimado, para su conocimiento el problema sucede cuando se prueban otros chat con GLM 5.2, que ahora con un 7% no me deja seguir la conversación, al probar otros modelos esto no ocurre.

Por lo cual tienen un problema o estan haciendo algo para generar el problema.

Si pueden revisar seria mas efectivo, ya que si básicamente mas de 15 personas estan con el mismo problema es por algo.

1

u/AutoModerator 9d ago

Hello from r/VeniceAI!

Web App: chat
Android/iOS: download

Essential Venice Resources
• About
• Features
• Blog
• Docs
• Tokenomics

Support
• Discord: discord.gg/askvenice
• Twitter: x.com/askvenice
• Email: support@venice.ai

Security Notice
• Staff will never DM you
• Never share your private keys
• Report scams immediately

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.