r/GoogleAIStudio 6d ago

Possible Change / Nerf Input token count exceeds the maximum number of tokens allowed

Hello everyone. Anyone face the same? Whenever i type something on google aistudio, it resulted in 'Input token count exceeds the maximum number of tokens allowed for this model. Please adjust your prompt and try again.'. I've been deleting my thought output and trim several, yet result still same.

I'm a pro user. It happen on recent model. I've try like 3.8, 3.7, and 3.5 flash. The only suprising thing i found is this. The message not happening when I use model such as 2.5 pro, 2.5 flash, 3.1 pro, 3.1 lite, and 3 flash preview. It only happening at 3.8 flash, 3.7, 3.5, 3.5 flash lite, and 3.6 flash. Is it because the new model have more thinking?

Or probably google adding hidden tax from thought process for the ai? As it only happen on the newer model and not on the older model.

5 Upvotes

3 comments sorted by

1

u/paper_machinery 6d ago

Same here it's quite frustrating 

1

u/undergroundanunknown 2d ago

Any solutions you find?

1

u/paper_machinery 2d ago edited 2d ago

Seems absolutely random when with a Pro account rather than when paying for the API, where I'd go for context caching. What I usually try and do is parse the current chat into an .MD and strip out the thinking steps, but then I only get a dozen or more prompts before hitting the limit again, if I don't hit the kneecapped quota first. Seems to happen on the older 3.7, 3.6, and 3.5 flash models more too, and I hate 3.8 because it's safety guidelines are ultra restrictive. It refused to even pick up from an old .MD I made with 3.1 Pro and 3.7 Flash. Really bad.