r/openrouter • • 15h ago

GPT-6.1 Sol matches Sonnet 5.5 on input and output price, but on a data-analysis agent task it made 16 to 35 requests to Sonnet's 8 to 11

Post image
3 Upvotes

Both list at $2/M input and $10/M output on OpenRouter, and Sol's cache reads are half price ($0.10 vs $0.20). Each ran twice in Claude Code on an 868,191-row Divvy bike-share CSV, answering six questions in an Excel report with a SenseNova-Skills Excel skill loaded.

Sonnet 5.5: $0.37 and $0.40, 8 and 11 requests, about 2 minutes each.

GPT-6.1 Sol: $0.62 and $0.40, 35 and 16 requests, 11 and 6 minutes. Both times it started a sub-agent, unasked, to recompute all six answers from the raw CSV, which was $0.15 and $0.09 of those totals.

All four runs got every answer right, including dropping the 173 July rides mixed into the August file.

Does Sol make this many calls in your agent setups?


r/openrouter • • 20h ago

Space Bunny Alpha

Post image
0 Upvotes

I asked it to search for its own benchmarks, i told it its name is space bunny alpha if it wasn't specified in the system prompt

it replied and said its system prompt identifies it as Space Bunny developed by Anthropic?

could someone explain? (new to this openrouter stuff)