scrapeunblocker-mcp-remote

Fetch any web page's HTML, AI-parsed JSON, or Google results via the ScrapeUnblocker anti-bot API

我该使用它吗

质量与安全性

A
描述质量
100%
模式完整度
100%
命名质量
95%
投毒风险
100%
权限匹配度
100%
协议合规性
100%

基于对工具定义和协议合规性的自动分析。

上下文开销

~1,277token 数(工具定义)
~2.6 KB典型响应大小
对注意力有中等影响(占 128k 上下文窗口的 1.00%)

这是每次将服务器的工具加载到模型上下文窗口时所消耗的大致 token 数。数值越高,可用于其他任务的注意力就越少。

安装

一键安装

将以下内容添加到你的 `claude_desktop_config.json` 文件中:

{
  "mcpServers": {
    "scrapeunblocker-mcp-remote": {
      "url": "https://mcp.scrapeunblocker.com/mcp"
    }
  }
}

远程端点

https://mcp.scrapeunblocker.com/mcpstreamable-http

它能做什么

工具清单

工具(4)

🟢 只读🟡 写入🔴 删除⚪ 未知
🟢fetch_html(url, proxy_country, wait_method, wait_value, sleep_seconds, ...)

Fetch the fully rendered HTML of any web page through the ScrapeUnblocker API (https://docs.scrapeunblocker.com), bypassing anti-bot protection (Cloudflare, DataDome, PerimeterX, Akamai, Shape). Use when a normal fetch is blocked (403/429, captcha) or the page needs a real browser. Returns raw HTML. Pass `steps` to interact with the page (search, click, paginate) before capture - use the list_elements tool first to discover selectors.

输入模式

{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "format": "uri",
      "description": "The absolute URL to fetch (http/https)."
    },
    "proxy_country": {
      "type": "string",
      "minLength": 2,
      "maxLength": 2,
      "description": "Optional ISO country code to route through, e.g. 'US'."
    },
    "wait_method": {
      "type": "string",
      "enum": [
        "css",
        "js"
      ],
      "description": "Optional render-wait: 'css' selector or 'js' expression."
    },
    "wait_value": {
      "type": "string",
      "description": "The selector/expression paired with wait_method."
    },
    "sleep_seconds": {
      "type": "number",
      "exclusiveMinimum": 0,
      "description": "Extra seconds to wait after load."
    },
    "steps": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "action": {
            "type": "string",
            "enum": [
              "wait_for",
              "wait_for_text",
              "wait",
              "click",
              "type",
              "select",
              "press_key",
              "scroll"
            ],
            "description": "The action to perform."
          },
          "selector": {
            "type": "string",
            "description": "CSS selector the action targets (required for wait_for/click/type/select)."
          },
          "selector_type": {
            "type": "string",
            "enum": [
              "css",
              "xPath",
              "className",
              "tagName"
            ],
            "description": "How to interpret `selector` (default 'css')."
          },
          "value": {
            "type": [
              "string",
              "number"
            ],
            "description": "Action payload: text to type/select, text for wait_for_text, a key name for press_key (e.g. 'Enter'), milliseconds for wait, or 'bottom'/pixels for scroll."
          },
          "clear": {
            "type": "boolean",
            "description": "For 'type': clear the field first."
          },
          "timeout_ms": {
            "type": "integer",
            "exclusiveMinimum": 0,
            "description": "Per-step timeout override in ms."
          }
        },
        "required": [
          "action"
        ],
        "additionalProperties": false
      },
      "description": "Ordered browser actions to run in a real browser after the page loads (wait_for, wait_for_text, wait, click, type [human-like], select, press_key, scroll), then return the resulting HTML. NON-IDEMPOTENT: it runs once and is not retried. A failed step returns a 422 naming the offending step plus the page HTML at that point. Discover selectors with list_elements first."
    }
  },
  "required": [
    "url"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢fetch_parsed(url, proxy_country)

Fetch a web page through the ScrapeUnblocker API (https://docs.scrapeunblocker.com) and return AI-parsed structured JSON instead of raw HTML (product details, article content, listings). If the page holds no structured data, the result says so (that call is not billed) - use fetch_html for the page itself.

输入模式

{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "format": "uri",
      "description": "The absolute URL to fetch and parse."
    },
    "proxy_country": {
      "type": "string",
      "minLength": 2,
      "maxLength": 2,
      "description": "Optional ISO country code, e.g. 'US'."
    }
  },
  "required": [
    "url"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢google_search(keyword, proxy_country, pages_to_check)

Run a Google search through the ScrapeUnblocker API (https://docs.scrapeunblocker.com) and return organic results as JSON.

输入模式

{
  "type": "object",
  "properties": {
    "keyword": {
      "type": "string",
      "minLength": 1,
      "description": "The search query."
    },
    "proxy_country": {
      "type": "string",
      "minLength": 2,
      "maxLength": 2,
      "description": "Optional ISO country code to search from, e.g. 'US'."
    },
    "pages_to_check": {
      "type": "integer",
      "exclusiveMinimum": 0,
      "maximum": 10,
      "description": "How many result pages to collect (default 1)."
    }
  },
  "required": [
    "keyword"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢list_elements(url, proxy_country, wait_method, wait_value, sleep_seconds)

Fetch a page through the ScrapeUnblocker API (https://docs.scrapeunblocker.com) and return its interactive elements (buttons, inputs, selects, links, forms), each with a ready-to-use selector, as JSON {url, count, elements:[...]} instead of raw HTML. Use it to discover what to target, then drive the page with the `steps` param of fetch_html.

输入模式

{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "format": "uri",
      "description": "The absolute URL to load and inspect (http/https)."
    },
    "proxy_country": {
      "type": "string",
      "minLength": 2,
      "maxLength": 2,
      "description": "Optional ISO country code to route through, e.g. 'US'."
    },
    "wait_method": {
      "type": "string",
      "enum": [
        "css",
        "js"
      ],
      "description": "Optional render-wait: 'css' selector or 'js' expression."
    },
    "wait_value": {
      "type": "string",
      "description": "The selector/expression paired with wait_method."
    },
    "sleep_seconds": {
      "type": "number",
      "exclusiveMinimum": 0,
      "description": "Extra seconds to wait after load before inspecting."
    }
  },
  "required": [
    "url"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}

社区

评价此服务器

证据

最近观测

已验证未记录版本4 个工具
已验证未记录版本4 个工具
已验证未记录版本4 个工具