scrapewright

Give it a URL, get structured rows. A model writes the parser once; replays are free.

사용해야 할까요

품질 및 안전성

B
설명 품질
90%
스키마 완전성
68%
이름 품질
80%
오염 위험
100%
권한 일치
100%
프로토콜 준수
100%

발견 사항 (1)

  • LOWTool 'account' description lacks action verbaccount에서

도구 정의와 프로토콜 준수에 대한 자동 분석을 기반으로 합니다.

컨텍스트 비용

~782토큰 (도구 정의)
~970 B일반적인 응답 크기
중간 정도의 주의 영향 (128k 컨텍스트의 0.61%)

이는 서버의 도구가 모델의 컨텍스트에 로드될 때마다 소비되는 대략적인 토큰 수입니다. 수치가 높을수록 다른 작업에 사용할 수 있는 주의가 줄어듭니다.

설치

원클릭 설치

`claude_desktop_config.json` 파일에 다음을 추가하세요:

{
  "mcpServers": {
    "scrapewright": {
      "url": "https://scrapewright.app/mcp"
    }
  }
}

원격 엔드포인트

https://scrapewright.app/mcpstreamable-http

할 수 있는 일

도구 목록

도구 (5)

🟢 읽기 전용🟡 쓰기🔴 삭제⚪ 알 수 없음
⚪detect_site(url)

Report what platform a site runs on and which strategy to use. Cheap; call it before a large job.

입력 스키마

{
  "type": "object",
  "properties": {
    "url": {
      "title": "Url",
      "type": "string"
    }
  },
  "required": [
    "url"
  ],
  "title": "detect_siteArguments"
}

출력 스키마

{
  "type": "object",
  "additionalProperties": true,
  "title": "detect_siteDictOutput"
}
🟢extract_page(url, fields, js, retry)

Extract structured data from ONE page. ``fields`` declares your own schema, e.g. ["title", "salary:number", "tags:list"]; omit it for the product schema. First call on a new site compiles a recipe (300 credits); later calls replay it for 1 credit per row. A site nobody could read is remembered for a week and answered without calling the model, so re-reading a dead page costs nothing. `retry` asks anyway.

입력 스키마

{
  "type": "object",
  "properties": {
    "url": {
      "title": "Url",
      "type": "string"
    },
    "fields": {
      "anyOf": [
        {
          "items": {
            "type": "string"
          },
          "type": "array"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Fields"
    },
    "js": {
      "default": false,
      "title": "Js",
      "type": "boolean"
    },
    "retry": {
      "default": false,
      "title": "Retry",
      "type": "boolean"
    }
  },
  "required": [
    "url"
  ],
  "title": "extract_pageArguments"
}

출력 스키마

{
  "type": "object",
  "additionalProperties": true,
  "title": "extract_pageDictOutput"
}
🟢crawl_site(listing_url, fields, max_items, js, scroll, ...)

Walk a site and extract every item. Waits up to four minutes; a longer crawl returns a job_id to pass to crawl_status. `mode` picks how the items are found. "rows" (the default) reads every card on the listing itself and follows its pagination -- right for search results, categories and tables. "links" follows each card into its own page, for sites where an item has one. "like" starts from one example item page, given as `like_page`, and finds every other page shaped like it; pass `listing_url` as well when you know the stock or results page, because inferring it is the part that fails.

입력 스키마

{
  "type": "object",
  "properties": {
    "listing_url": {
      "title": "Listing Url",
      "type": "string"
    },
    "fields": {
      "anyOf": [
        {
          "items": {
            "type": "string"
          },
          "type": "array"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Fields"
    },
    "max_items": {
      "default": 25,
      "title": "Max Items",
      "type": "integer"
    },
    "js": {
      "default": false,
      "title": "Js",
      "type": "boolean"
    },
    "scroll": {
      "default": 0,
      "title": "Scroll",
      "type": "integer"
    },
    "mode": {
      "default": "rows",
      "title": "Mode",
      "type": "string"
    },
    "like_page": {
      "anyOf": [
        {
          "type": "string"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Like Page"
    }
  },
  "required": [
    "listing_url"
  ],
  "title": "crawl_siteArguments"
}

출력 스키마

{
  "type": "object",
  "additionalProperties": true,
  "title": "crawl_siteDictOutput"
}
🟢crawl_status(job_id)

Fetch a crawl that outlived its call.

입력 스키마

{
  "type": "object",
  "properties": {
    "job_id": {
      "title": "Job Id",
      "type": "string"
    }
  },
  "required": [
    "job_id"
  ],
  "title": "crawl_statusArguments"
}

출력 스키마

{
  "type": "object",
  "additionalProperties": true,
  "title": "crawl_statusDictOutput"
}
⚪account

Credits left and this month's usage for the key in use.

입력 스키마

{
  "type": "object",
  "properties": {},
  "title": "accountArguments"
}

출력 스키마

{
  "type": "object",
  "additionalProperties": true,
  "title": "accountDictOutput"
}

커뮤니티

이 서버 평가하기

증거

최근 관측

검증됨버전이 기록되지 않음도구 5개
검증됨버전이 기록되지 않음도구 5개
검증됨버전이 기록되지 않음도구 5개