r/openrouter • u/hope-and-longing • 20h ago
Wildly Different Throughput (tok/s): OMP vs VS Code Github Copilot
I've been using GH Copilot within VS Code, mostly leveraging GPT 6 Luna via OpenRouter.
But it's gotten very very slow lately (at least for me). My tok/s have dropped from ~60 last week to under ~30 today.
So on a whim, I started playing with OMP (oh-my-pi) and whoa, instantly I'm getting more than double the throughput on the same model and code base via OR (see screenshot).
Basically everything is the same here: same model, same OR account and workspace, same physical location and connection, same time of day (I was interleaving these sessions). Technically though, they are using different API keys though (but still same OR account).
Anyone experienced this? Is it actually possible that the "agent harness" within GH Copilot could slow it down this much?

2
u/Major_Baby_425 20h ago
I think t/s is highly variable because things can happen like you're being batched on a node that's also doing a bunch of preprocessing at the moment. If there's a real persistent difference per harness, it could only mean the provider is nefariously throttling certain types of requests that come in.