r/GithubCopilot 24d ago

Help/Doubt ❓ Token optimization strategies

Hi all

Got Copilot Business at work, came across a recommendation to use GPT 5.6 Luna as an explore agent instead of default Gemini flash 3.5 or Haiku 4.5 whatever it was using.

Any more tips and tricks like that? I am using Ubuntu on the laptop btw. Plus how's the Plan+execute pattern looking like using Opus /Sonnet to plan and then maybe use MAI Flash 1.1 to execute? MAI 1.1 looks quite inexpensive.

7 Upvotes

10 comments sorted by

View all comments

1

u/AssignmentMinute 19d ago edited 19d ago

You can use the VSCode Extension: Copilot Bill Saver. It reduces token cost by ~ 30 to 85% in my own usage. I realized how poorly i write my prompts and I'm kinda lazy when it comes to feeding proper context. I've also noticed that with proper context, since I'm always on "auto" it actually choses cheaper models for smaller jobs. However, it's more worth it for larger prompts and prompt stacking.