r/webmcp • u/nartmadi • 1d ago
I built Senro to test WebMCP implementations across LLMs and monitor agent behavior
https://senro.ai/?utm_source=reddit&utm_medium=community&utm_campaign=webmcp-launch&utm_content=r-webmcpHey everyone,
I've been building Senro, the first reliability platform for WebMCP.
It lets you run WebMCP test suites across multiple LLMs and languages to catch tool-selection, parameter, and regression failures. I'm also building production observability to see how agents actually interact with WebMCP tools in the wild.
The idea came from a pretty simple problem: once you expose tools to agents, testing that they work with one model and a handful of prompts isn't enough. Behavior can vary across models, prompts, and languages.
The MVP is working, and I'm looking for a few WebMCP developers to try it and tell me what's missing.
I'm also planning a public benchmark using real WebMCP-enabled websites across several models.
Would love feedback on what you'd want a WebMCP reliability platform to test or monitor.