scrapewright
Give it a URL, get structured rows. A model writes the parser once; replays are free.
Sollte ich dies verwenden
Qualität und Sicherheit
Befunde (1)
- LOWin account
Basierend auf einer automatisierten Analyse der Tool-Definitionen und der Einhaltung des Protokolls.
Kontextkosten
Dies ist die ungefähre Anzahl der Tokens, die jedes Mal verbraucht werden, wenn die Tools des Servers in den Kontext eines Modells geladen werden. Höhere Werte verringern die Aufmerksamkeit, die für andere Aufgaben verfügbar ist.
Installieren
Installation mit einem Klick
Fügen Sie dies Ihrer Datei `claude_desktop_config.json` hinzu:
{
"mcpServers": {
"scrapewright": {
"url": "https://scrapewright.app/mcp"
}
}
}Remote-Endpunkte
https://scrapewright.app/mcpstreamable-httpWas es kann
Tool-Inventar
Tools (5)
⚪detect_site(url)
Report what platform a site runs on and which strategy to use. Cheap; call it before a large job.
Eingabe-Schema
{
"type": "object",
"properties": {
"url": {
"title": "Url",
"type": "string"
}
},
"required": [
"url"
],
"title": "detect_siteArguments"
}Ausgabe-Schema
{
"type": "object",
"additionalProperties": true,
"title": "detect_siteDictOutput"
}🟢extract_page(url, fields, js)
Extract structured data from ONE page. ``fields`` declares your own schema, e.g. ["title", "salary:number", "tags:list"]; omit it for the product schema. First call on a new site compiles a recipe (300 credits); later calls replay it for 1 credit per row.
Eingabe-Schema
{
"type": "object",
"properties": {
"url": {
"title": "Url",
"type": "string"
},
"fields": {
"anyOf": [
{
"items": {
"type": "string"
},
"type": "array"
},
{
"type": "null"
}
],
"default": null,
"title": "Fields"
},
"js": {
"default": false,
"title": "Js",
"type": "boolean"
}
},
"required": [
"url"
],
"title": "extract_pageArguments"
}Ausgabe-Schema
{
"type": "object",
"additionalProperties": true,
"title": "extract_pageDictOutput"
}🟢crawl_site(listing_url, fields, max_items, js, scroll)
Walk a site from one listing URL and extract every item. Waits up to four minutes; a longer crawl returns a job_id to pass to crawl_status.
Eingabe-Schema
{
"type": "object",
"properties": {
"listing_url": {
"title": "Listing Url",
"type": "string"
},
"fields": {
"anyOf": [
{
"items": {
"type": "string"
},
"type": "array"
},
{
"type": "null"
}
],
"default": null,
"title": "Fields"
},
"max_items": {
"default": 25,
"title": "Max Items",
"type": "integer"
},
"js": {
"default": false,
"title": "Js",
"type": "boolean"
},
"scroll": {
"default": 0,
"title": "Scroll",
"type": "integer"
}
},
"required": [
"listing_url"
],
"title": "crawl_siteArguments"
}Ausgabe-Schema
{
"type": "object",
"additionalProperties": true,
"title": "crawl_siteDictOutput"
}🟢crawl_status(job_id)
Fetch a crawl that outlived its call.
Eingabe-Schema
{
"type": "object",
"properties": {
"job_id": {
"title": "Job Id",
"type": "string"
}
},
"required": [
"job_id"
],
"title": "crawl_statusArguments"
}Ausgabe-Schema
{
"type": "object",
"additionalProperties": true,
"title": "crawl_statusDictOutput"
}⚪account
Credits left and this month's usage for the key in use.
Eingabe-Schema
{
"type": "object",
"properties": {},
"title": "accountArguments"
}Ausgabe-Schema
{
"type": "object",
"additionalProperties": true,
"title": "accountDictOutput"
}Community
Nachweis