RunBoth

Ask runboth.dev in plain language: what RunBoth does, how to install it, and its limits.

Should I use this

Quality & Safety

A
Description quality
93%
Schema completeness
77%
Naming quality
87%
Poisoning risk
100%
Permission match
100%
Protocol compliance
100%

Based on automated analysis of tool definitions and protocol compliance.

Context Cost

~456Tokens (tool definitions)
~791 BTypical response size
Minimal attention impact (0.36% of 128k context)

This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.

Install

One-Click Install

Add this to your `claude_desktop_config.json` file:

{
  "mcpServers": {
    "agent": {
      "url": "https://runboth-agent-mgbxvs4loq-uc.a.run.app/mcp"
    }
  }
}

Remote endpoints

https://runboth-agent-mgbxvs4loq-uc.a.run.app/mcpstreamable-http

What it can do

Tool inventory

Tools (3)

🟢 Read-only🟡 Write🔴 Delete⚪ Unknown
🟡ask(query, site, generate_mode)

Ask runboth.dev a question in plain language and get the answer from its own pages, with sources. RunBoth is a zero-dependency Python tool that finds out what a refactor actually changed by running the old and the new code on generated inputs and comparing seven behaviour channels, rather than reading the diff. Covers: detecting behavioural change from a Python refactor; differential execution and the seven channels it compares; installing RunBoth and running it as a commit hook or GitHub Action; what a CHANGED, NO CHANGE or ABSTAINED verdict means; how RunBoth was red-teamed and where its limits are. For example: How do I find out what my Python refactor actually changed in behaviour? Answers come back as ranked pages from runboth.dev with a URL, a score and a one-line reason each.

Input Schema

{
  "type": "object",
  "properties": {
    "query": {
      "type": "string",
      "description": "The question, in plain language. For example: How do I find out what my Python refactor actually changed in behaviour?"
    },
    "site": {
      "type": "array",
      "items": {
        "type": "string"
      },
      "description": "Optional list of sites to search. If not provided, searches all configured sites"
    },
    "generate_mode": {
      "type": "string",
      "enum": [
        "list",
        "generate",
        "summarize"
      ],
      "description": "The type of response to generate",
      "default": "list"
    }
  },
  "required": [
    "query"
  ]
}
🟢list_sites

List the sites this endpoint answers for (runboth.dev).

Input Schema

{
  "type": "object",
  "properties": {}
}
🟢trust_lookup(target, top_k)

Look up what AI agents recorded after calling a website or an MCP endpoint: which capability they used, whether it worked, how long it took, and what they learned. Read-only, no key. Use it before you trust an endpoint you have never called, including this one.

Input Schema

{
  "type": "object",
  "properties": {
    "target": {
      "type": "string",
      "description": "A domain or endpoint URL, for example runboth.dev or https://host/mcp"
    },
    "top_k": {
      "type": "integer",
      "description": "How many records to return, 1 to 25",
      "default": 8
    }
  },
  "required": [
    "target"
  ]
}

Community

Rate this Server

Evidence

Recent observations

verifiedversion not recorded3 tools
verifiedversion not recorded3 tools