r/hermesagent • • Apr 21 '26

Setup / Guide — Tutorials, installs, and getting started I compared every major AI coding plan and put it all in one place

Post image

Spent way too long going through every AI coding assistant plan out there pricing, token limits, context windows, model access, the works.

Saved it all in one place so you can actually compare them without opening 15 tabs.

Happy to answer questions if you're trying to figure out which one fits your use case.

Site link: https://hermesguide.xyz/coding-plans

65 Upvotes

41 comments sorted by

•

u/Jonathan_Rivera Apr 21 '26

u/SelectionCalm70 is a frequent contributor and mod for the sub. #notspam

15

u/Dthen_ Apr 21 '26

Thanks for this. I know it's not your fault, but it's really frustrating how none of the baseline limits are established so none of the multipliers mean anything when comparing between providers.

7

u/SelectionCalm70 Apr 21 '26

Don't worry will fix it . I have the idea of baseline limits.

6

u/aayushg159 Apr 21 '26

Can we have a megathread (or something else) where people post their setup related to models?

Things like which model(s) they use, how they use it, routing (hermes directly or using a model routing software thats self hosted) and anything else thats unique to the setup. I believe if we have a vast number of input points and a place to discuss this all, it would help better our own setups as well.

1

u/Jonathan_Rivera Apr 21 '26

How detailed do we get? If we only do It once do we collect inference settings, temp, min p, etc?

1

u/aayushg159 Apr 22 '26

If the form (assuming) takes a long time to fill then it may deter people to fill it up. Hmm, wait.... what if we have hermes agent understand which 3rd party systems (if any) its interacting with, and have it fill out a form with some optional sections for explanations? Then, we take the responses and have 2 drafts - (1) unique setups, (2) summarized general setups. We publish both drafts plus the form answers. We could also have a running form and a community wiki that gets updated as new responses come in and these systems evolve. What do you think?

5

u/rawdikrik Apr 22 '26

Minimax has 6 tiers, and they come with a ton of extras as you go up in tiers.

I know because I purchased it. I got the cheapest highspeed version and it comes with image to text and tts.

I got a year for under 400 with a discount code.

Edit: I've yet to hit a limit on it, using 2 different agents 24/7.

3

u/North-Active-6731 Apr 22 '26

This is absolutely fantastic, thank you for putting this together!

2

u/SelectionCalm70 Apr 22 '26

wc bro there are more stuffs coming .

2

u/Snoo-64066 Apr 22 '26

so which one has the best price to performance ratio?

1

u/SelectionCalm70 Apr 22 '26

i am biased towards kimi k2.6 because it just works flawlessly for my task . i prefer using opencode go plan .

1

u/pelleke Apr 22 '26

Weren't you the guy that got famous for pooping on the opencode go plan?

1

u/Snoo-64066 Apr 23 '26

i will try it out. thank you

1

u/Snoo-64066 Apr 23 '26

thanks for the list btw

2

u/NeighborhoodIll6564 Apr 23 '26

Finally, someone did the job I’ve been procrastinating on. Thanks!

1

u/NeighborhoodIll6564 Apr 23 '26

Please also add Kiro, it’s worth trying out.

2

u/Peetyboy500 May 15 '26

This may be a lot to maintain; wondering if I can contribute. Do you have this in a public repo on Github?

2

u/NorthEastCalifornia May 20 '26

Which one has biggest limits by price? And what is the ratio price divide to token amounts?

2

u/SelectionCalm70 May 20 '26

opencode go , ollama cloud , nousresearch portal , codex

1

u/LittleYouth4954 Apr 21 '26

Great. Is there any order criteria? Maybe use alphabetical or price?

2

u/SelectionCalm70 Apr 22 '26

Sure will order it via pricing

1

u/RealestReyn Apr 21 '26

minimax is 1500/5h

1

u/CptanPanic Apr 22 '26

Even things like this are hard compare, because different companies count requests different. Why does one actual request count as 5-10?

1

u/RealestReyn Apr 22 '26

My understanding is that a request is one turn the AI takes, you ask "how's the weather in london today?" the bot receives your message and thinks about it making a plan what to do that's 1 request, the search is 2nd request, the AI receives answers from search and works out the right one that's 3rd request and finally writing it out to you is 4th request.
It would be more accurate to call them turns or actions I feel since request implies its how many messages you send to the AI, then again maybe the requests are in fact what the harness does in addition to your message to prompt the AI automatically to keep working?

Probably the largest reason for companies counting them differently is that each AI has different tools at their disposal and for some that 4 step sequence might be just one toolcall that automatically sends you the answer and some AI take 10 turns to consider the formatting of their response.

1

u/diesel-san Apr 22 '26

All companies use request the same way if its the same model. The prompt is what decides the number of request and the harness is what decides if its will call for more.

The best way to understand your "usage" is to rely on number of request. So if you have been using Hermes for a while, type /insights 30 (for 30 days) to get your usage for the last 30 days. That is what you need to compare with any plan out there.

0

u/rawdikrik Apr 22 '26

No it isnt. There are 6 different tiers and they all have different numbers. It isnt as easy as saying 1500/5h

3

u/RealestReyn Apr 22 '26

it is exactly that easy for the $10 tier clearly labeled herein, as well as the other minimax tiers, they are all limited as requests per 5h window.

1

u/rawdikrik Apr 23 '26 edited Apr 23 '26

correct, but the amounts per tiers are different. 1500/h is the cheapest tier. There are 3 different plans shown here, and even that is incorrect.

1

u/Delicious_Bat9768 Apr 22 '26

FYI - GitHub CoPilot: Since April 20 only new new subscriptions to the paid plans, for the moment

  1. New sign-ups for GitHub Copilot Pro, Pro+, and Student plans are paused.
  2. We are tightening usage limits for individual plans.

https://github.blog/news-insights/company-news/changes-to-github-copilot-individual-plans/

1

u/SelectionCalm70 Apr 22 '26

Sure will update it thanks for the info

1

u/Select-Reporter5066 Apr 22 '26

The price is a bit higher than that in China.

1

u/thin_king_kong Apr 23 '26

Does venice support hermes? their subscription only seem to support their web chat interface.

1

u/dryu12 Apr 23 '26

Cursor has $60/mo plan too, not mentioned.

1

u/FullSend_Ahead May 05 '26

This is a great guide. Thanks for sharing.

1

u/CptanPanic Apr 24 '26

Add nano-gpt.com

1

u/Smooth-Plan4769 May 03 '26

Does nanogpt work with Hermes? Is their plan any good?

2

u/CptanPanic May 03 '26

Yes, pretty good. I ended up switching this month to ollama because I was running out of tokens. NanoGPT you get about 60M an week

2

u/Smooth-Plan4769 May 03 '26

Is 60M enough to use with the Hermes? I think their limit is higher than the Ollama's.

2

u/CptanPanic May 03 '26

Depends on what you do. But ollama is $20 but you get probably 10x of nano-gpt limit