scrapewright

Give it a URL, get structured rows. A model writes the parser once; replays are free.

我該用這個嗎

品質與安全性

B
說明品質
90%
結構描述完整度
68%
命名品質
80%
汙染風險
100%
權限相符程度
100%
協定合規性
100%

發現項目(1)

  • LOWTool 'account' description lacks action verb在 account 中

根據工具定義與協定合規性的自動化分析。

上下文成本

~567Token(工具定義)
~844 B典型回應大小
極小的注意力影響(128k 上下文的 0.44%)

這是每次將伺服器的工具載入模型上下文時所消耗的約略 token 數量。數量越高,可用於其他工作的注意力就越少。

安裝

一鍵安裝

將以下內容加入你的 `claude_desktop_config.json` 檔案:

{
  "mcpServers": {
    "scrapewright": {
      "url": "https://scrapewright.app/mcp"
    }
  }
}

遠端端點

https://scrapewright.app/mcpstreamable-http

它能做什麼

工具清單

工具(5)

🟢 唯讀🟡 寫入🔴 刪除⚪ 未知
⚪detect_site(url)

Report what platform a site runs on and which strategy to use. Cheap; call it before a large job.

輸入結構描述

{
  "type": "object",
  "properties": {
    "url": {
      "title": "Url",
      "type": "string"
    }
  },
  "required": [
    "url"
  ],
  "title": "detect_siteArguments"
}

輸出結構描述

{
  "type": "object",
  "additionalProperties": true,
  "title": "detect_siteDictOutput"
}
🟢extract_page(url, fields, js)

Extract structured data from ONE page. ``fields`` declares your own schema, e.g. ["title", "salary:number", "tags:list"]; omit it for the product schema. First call on a new site compiles a recipe (300 credits); later calls replay it for 1 credit per row.

輸入結構描述

{
  "type": "object",
  "properties": {
    "url": {
      "title": "Url",
      "type": "string"
    },
    "fields": {
      "anyOf": [
        {
          "items": {
            "type": "string"
          },
          "type": "array"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Fields"
    },
    "js": {
      "default": false,
      "title": "Js",
      "type": "boolean"
    }
  },
  "required": [
    "url"
  ],
  "title": "extract_pageArguments"
}

輸出結構描述

{
  "type": "object",
  "additionalProperties": true,
  "title": "extract_pageDictOutput"
}
🟢crawl_site(listing_url, fields, max_items, js, scroll)

Walk a site from one listing URL and extract every item. Waits up to four minutes; a longer crawl returns a job_id to pass to crawl_status.

輸入結構描述

{
  "type": "object",
  "properties": {
    "listing_url": {
      "title": "Listing Url",
      "type": "string"
    },
    "fields": {
      "anyOf": [
        {
          "items": {
            "type": "string"
          },
          "type": "array"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Fields"
    },
    "max_items": {
      "default": 25,
      "title": "Max Items",
      "type": "integer"
    },
    "js": {
      "default": false,
      "title": "Js",
      "type": "boolean"
    },
    "scroll": {
      "default": 0,
      "title": "Scroll",
      "type": "integer"
    }
  },
  "required": [
    "listing_url"
  ],
  "title": "crawl_siteArguments"
}

輸出結構描述

{
  "type": "object",
  "additionalProperties": true,
  "title": "crawl_siteDictOutput"
}
🟢crawl_status(job_id)

Fetch a crawl that outlived its call.

輸入結構描述

{
  "type": "object",
  "properties": {
    "job_id": {
      "title": "Job Id",
      "type": "string"
    }
  },
  "required": [
    "job_id"
  ],
  "title": "crawl_statusArguments"
}

輸出結構描述

{
  "type": "object",
  "additionalProperties": true,
  "title": "crawl_statusDictOutput"
}
⚪account

Credits left and this month's usage for the key in use.

輸入結構描述

{
  "type": "object",
  "properties": {},
  "title": "accountArguments"
}

輸出結構描述

{
  "type": "object",
  "additionalProperties": true,
  "title": "accountDictOutput"
}

社群

為此伺服器評分

證據

近期觀測

已驗證未記錄版本5 個工具
已驗證未記錄版本5 個工具