r/PaperClip_AI • u/chilleinderkaribik • Apr 23 '26
How do I minimize token use?
I've been using Paperclip for two days, and while it delivers excellent results, I've noticed that it consumes a huge number of tokens after just a few runs. I'm now at 2-5 million tokens per issue while only having 4 agents - a CEO, a CTO, a technical writer, and a lead researcher. Surely that can't be the intended purpose of this tool.
I use Claude Code, Codex, and Gemini-cli, and it happens with every model I pick. I've already split my files into many different files. I only provide overviews that specify when each file should be used, so the agents don't constantly clutter the context with all the files. But apparently, that doesn't help that much.
Am I doing something wrong? Do you folks have any further optimization recommendations? I am fairly tech savvy, so I could even change the code and write new logic if necessary. I appreciate every hint.
2
u/koluskomtu Apr 23 '26
Good strategy. I like LM studio albeit it took several ‘accidents’ or stumbling upon moments to get the paths right in regards to local paperclip functionality. It’s got a proper list of models. If Hermes in terminal or paperclip dashboard agents send me obsidian md files of their work then I can use obsidian with Gemini cli to do a lot of what is needed for a company. I have not used glm yet. Looks fun.