Vouch
Vouch — independently measured reliability scores for MCP tools, not self-reported claims.
Should I use this
Quality & Safety
Based on automated analysis of tool definitions and protocol compliance.
Context Cost
This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.
Install
One-Click Install
Add this to your `claude_desktop_config.json` file:
{
"mcpServers": {
"vouchtools": {
"url": "https://vouch.tools/api/mcp"
}
}
}Remote endpoints
https://vouch.tools/api/mcpstreamable-httpWhat it can do
Tool inventory
Tools (3)
🟢vouch_check(tool)
Get an MCP tool's independently measured trust score before relying on it: real invocation trials, not self-reported metadata. Call this before adding a new MCP server to a project, or when deciding between two tools that do similar things. Look up by tool name, package, or server — free text.
Input Schema
{
"type": "object",
"properties": {
"tool": {
"type": "string",
"description": "A tool name, 'server tool', or package name — free text."
}
},
"required": [
"tool"
],
"$schema": "https://json-schema.org/draft/2020-12/schema"
}🟢vouch_find(query, tier, limit)
Search Vouch's independently measured corpus of MCP tools by free-text query. Call this when choosing an MCP tool for a task and you want options ranked by measured reliability rather than popularity — optionally filtered by measurement tier (deep = full behavioural battery, shallow = schema + sample-call only).
Input Schema
{
"type": "object",
"properties": {
"query": {
"type": "string",
"description": "Free-text search — matches tool name, description, or server name."
},
"tier": {
"description": "Restrict to tools measured at this tier.",
"type": "string",
"enum": [
"shallow",
"deep"
]
},
"limit": {
"type": "integer",
"minimum": 1,
"maximum": 50
}
},
"required": [
"query"
],
"$schema": "https://json-schema.org/draft/2020-12/schema"
}⚪vouch_compare(tools)
Compare 2-5 MCP tools' measured behaviour side by side before picking one. Call this when weighing alternatives for the same task. Never collapses the comparison into a single ranked number — each tool keeps its own score and tier, and a caveat is surfaced when compared tools weren't measured at the same tier.
Input Schema
{
"type": "object",
"properties": {
"tools": {
"minItems": 2,
"maxItems": 5,
"type": "array",
"items": {
"type": "string"
},
"description": "Tool names, packages, or servers — free text, one per tool."
}
},
"required": [
"tools"
],
"$schema": "https://json-schema.org/draft/2020-12/schema"
}Community
Evidence