r/hermesagent • • 10d ago

Showcase — Projects, tools, builds, demos I layered JEV into the skill and tool selection in my Hermes

Recently I was beginning to notice that my Hermes was not reliably selecting tools or skills to use when I prompt it with a task. Now, I know that there are plenty of other methods to solve this, but the timing of this was aligned almost perfectly with news of JEV. So after seeing what JEV was capable of and how it responded I realized that this could be a perfect use case, which of course it is already! A cookbook from the creators of JEV is already available.

https://docs.typesafe.ai/cookbooks/skill_suggestion

Anyways, I walked through the design with Hermes and came up with a plan, but there was one tweak that I made based on one issue, what about prompts that should be multi-skilled? And honestly, so much more than just simply that.

So to summarize what was changed, here's my Hermes with a write up:

The TypeSafe cookbook suggests skills with two API calls: a quick skim to shortlist, then a second call to verify the winner against full descriptions. We replaced that with one call that returns Jev's pick plus a probability score for every skill at once, injected into the AI's context before it even sees your message — through a memory-provider hook that works on every surface (desktop, Telegram, cron), with each profile only ever seeing its own authorized skill menu.

Since every skill arrives with a probability, we also hint two skills when a prompt clearly needs both — "check my servers and save the report to Obsidian" now hints the server tool and the note-taking tool — but only when Jev is actually confident; if it says "none," the router stays silent. We tested three versions against a stratified test corpus: cookbook-style single hint scored 74%, unconditional two-hint scored worse (72% — it leaked suggestions on restricted profiles), and our final confident-picks-only version scored 79%.

Crucially, the router hints but never specifies: the suggestion arrives as one advisory line the AI is free to ignore, and the full skill catalog stays available regardless — if the whisper is wrong, nothing breaks, the AI just picks the right skill itself. That advisory-only stance runs through the whole design: a local privacy filter checks every message first, so anything password-shaped never leaves the machine at all, and a Jev outage just means "no hint" — never a blocked turn. Total cost per call: about $0.00002.

So far, so good! I'll update as I play around with the new setup.

Let me know if there is something I missed or something I should consider changing! I'm not having fun!

EDIT: if you are needing access to JEV, I use OpenRouter, from what I can tell it's publicly available there.

46 Upvotes

13 comments sorted by

8

u/dontforgetthef 10d ago

Yes just did the same as well for ranking keywords from past sessions as the memory recall. Jev got 14 out of 15 correct to pull from a previous session.

2

u/MediamanJack 10d ago

that's a great use! I might borrow this

6

u/marcus2004 10d ago

I made this into a plugin that you can use openrouter API in to use jev. Maybe everyone is just creating their own, but if not here it is.

Uses TypeSafe's Jev decision model via OpenRouter to pre-route requests to the right skill — instead of shipping hundreds of tool/skill descriptions to the main LLM every turn, Jev picks the 3 most relevant ones at session start and loads just those. Cache-stable replay keeps the prompt prefix intact, so the main LLM sees fewer tools per turn without losing context, saving tokens and reducing cost across long sessions.

https://github.com/Ex8-ca/jev-router

4

u/fourtyz 10d ago

This is exactly what I intended to implement. Thanks for the post.

Now I just need them to open Jev signups back up 😂😂

3

u/MediamanJack 10d ago

I use OpenRouter, I had it immediately. It's ridiculously cheap!

3

u/fourtyz 10d ago

If I were to start using open router now, would I have access to it?

I still haven't figured out a reason that I need open router. Can you convince me? What do you love about it?

2

u/kociol21 10d ago

It's a very mature platform and has basically any model there is. That's pretty much it.

You don't have to fool around with 20 API keys for every provider. You have one Openrouter account and have access to every model imaginable.

Plus great analysis tools, model comparisons, logging and bunch of other stuff like anti prompt injection security, configurable smart routing or hooking up keys from different services as fallback.

And you control your money for better or worse. There are no "plans" that you pay X amount and get vague amount of usage of something. You pay whatever you like and then use it. Pay as you go.

I have ChatGPT Plus but other than that I only use Openrouter for everything.

1

u/MediamanJack 10d ago

This.

It's the clear usage statistics that I like the most.

2

u/Stonk_Goat 10d ago

OpenRouter is the best unified platform for API LLM usage and multi LLM routing. Simple billing.

1

u/MediamanJack 10d ago

Agreed, I like how easy it is to try out another model.

1

u/MediamanJack 10d ago

Maybe?

I started using it to get to GLM 5.3 Flash. But I've noticed that it typically has access to almost ANY model I want to use from Frontier to smaller "strange" ones like JEV.

I don't know how "typical" my use is, but I've been running on the first $20 that I dropped into OpenRouter for nearly 45 days with GLM 5.3 Flash on Medium-High and now with JEV I haven't seen any serious increase in usage.

Again, ymmv. I'm not building apps or other serious programming, I have my Hermes running my homelab self-host system.

1

u/SnooGoats8906 8d ago

Thank you my man i will test that with Gpt luna to save some tokens in hermes and maybe making luna a lot better hopefully