scrapewright
Give it a URL, get structured rows. A model writes the parser once; replays are free.
我該用這個嗎
品質與安全性
發現項目(1)
- LOW在 account 中
根據工具定義與協定合規性的自動化分析。
上下文成本
這是每次將伺服器的工具載入模型上下文時所消耗的約略 token 數量。數量越高,可用於其他工作的注意力就越少。
安裝
一鍵安裝
將以下內容加入你的 `claude_desktop_config.json` 檔案:
{
"mcpServers": {
"scrapewright": {
"url": "https://scrapewright.app/mcp"
}
}
}遠端端點
https://scrapewright.app/mcpstreamable-http它能做什麼
工具清單
工具(5)
⚪detect_site(url)
Report what platform a site runs on and which strategy to use. Cheap; call it before a large job.
輸入結構描述
{
"type": "object",
"properties": {
"url": {
"title": "Url",
"type": "string"
}
},
"required": [
"url"
],
"title": "detect_siteArguments"
}輸出結構描述
{
"type": "object",
"additionalProperties": true,
"title": "detect_siteDictOutput"
}🟢extract_page(url, fields, js)
Extract structured data from ONE page. ``fields`` declares your own schema, e.g. ["title", "salary:number", "tags:list"]; omit it for the product schema. First call on a new site compiles a recipe (300 credits); later calls replay it for 1 credit per row.
輸入結構描述
{
"type": "object",
"properties": {
"url": {
"title": "Url",
"type": "string"
},
"fields": {
"anyOf": [
{
"items": {
"type": "string"
},
"type": "array"
},
{
"type": "null"
}
],
"default": null,
"title": "Fields"
},
"js": {
"default": false,
"title": "Js",
"type": "boolean"
}
},
"required": [
"url"
],
"title": "extract_pageArguments"
}輸出結構描述
{
"type": "object",
"additionalProperties": true,
"title": "extract_pageDictOutput"
}🟢crawl_site(listing_url, fields, max_items, js, scroll)
Walk a site from one listing URL and extract every item. Waits up to four minutes; a longer crawl returns a job_id to pass to crawl_status.
輸入結構描述
{
"type": "object",
"properties": {
"listing_url": {
"title": "Listing Url",
"type": "string"
},
"fields": {
"anyOf": [
{
"items": {
"type": "string"
},
"type": "array"
},
{
"type": "null"
}
],
"default": null,
"title": "Fields"
},
"max_items": {
"default": 25,
"title": "Max Items",
"type": "integer"
},
"js": {
"default": false,
"title": "Js",
"type": "boolean"
},
"scroll": {
"default": 0,
"title": "Scroll",
"type": "integer"
}
},
"required": [
"listing_url"
],
"title": "crawl_siteArguments"
}輸出結構描述
{
"type": "object",
"additionalProperties": true,
"title": "crawl_siteDictOutput"
}🟢crawl_status(job_id)
Fetch a crawl that outlived its call.
輸入結構描述
{
"type": "object",
"properties": {
"job_id": {
"title": "Job Id",
"type": "string"
}
},
"required": [
"job_id"
],
"title": "crawl_statusArguments"
}輸出結構描述
{
"type": "object",
"additionalProperties": true,
"title": "crawl_statusDictOutput"
}⚪account
Credits left and this month's usage for the key in use.
輸入結構描述
{
"type": "object",
"properties": {},
"title": "accountArguments"
}輸出結構描述
{
"type": "object",
"additionalProperties": true,
"title": "accountDictOutput"
}社群
證據