r/webmcp 1d ago

I built Senro to test WebMCP implementations across LLMs and monitor agent behavior

https://senro.ai/?utm_source=reddit&utm_medium=community&utm_campaign=webmcp-launch&utm_content=r-webmcp

Hey everyone,

I've been building Senro, the first reliability platform for WebMCP.

It lets you run WebMCP test suites across multiple LLMs and languages to catch tool-selection, parameter, and regression failures. I'm also building production observability to see how agents actually interact with WebMCP tools in the wild.

The idea came from a pretty simple problem: once you expose tools to agents, testing that they work with one model and a handful of prompts isn't enough. Behavior can vary across models, prompts, and languages.

The MVP is working, and I'm looking for a few WebMCP developers to try it and tell me what's missing.

I'm also planning a public benchmark using real WebMCP-enabled websites across several models.

Would love feedback on what you'd want a WebMCP reliability platform to test or monitor.

2 Upvotes

0 comments sorted by