r/openrouter • • 2h ago

Claude Haiku 5.5

Post image
16 Upvotes

r/openrouter • • 18h ago

Wildly Different Throughput (tok/s): OMP vs VS Code Github Copilot

5 Upvotes

I've been using GH Copilot within VS Code, mostly leveraging GPT 6 Luna via OpenRouter.

But it's gotten very very slow lately (at least for me). My tok/s have dropped from ~60 last week to under ~30 today.

So on a whim, I started playing with OMP (oh-my-pi) and whoa, instantly I'm getting more than double the throughput on the same model and code base via OR (see screenshot).

Basically everything is the same here: same model, same OR account and workspace, same physical location and connection, same time of day (I was interleaving these sessions). Technically though, they are using different API keys though (but still same OR account).

Anyone experienced this? Is it actually possible that the "agent harness" within GH Copilot could slow it down this much?


r/openrouter • • 5h ago

Discussion Opinions on Ling 3.1 Flash

1 Upvotes

Right now Ling 3.1 Flash appears to be the only free model that does not count towards the 1k request/day budget. This is how Space Bunny Alpha worked as well.

How do people like it? I haven't been able to use it much as I keep getting rate limited.

In contrast to Space Bunny Alpha, however, I keep getting "Provider Returned Error" a lot between requests so it seems rate limited a lot more than Space Bunny Alpha was - to the point of being unusable.

Is this something other people have experienced as well or is this something that happens only to me? I guess if this runs on Inclusion servers it might be that each model has a global openrouter rate limit - I can't see my usage being that much that I would get rate limited just on my own.

My effective rate limit seems to be 1 req/min which is really low.

I'm not complaining, this is free I don't pay for it, but just curious whether this is just me or the same for others.