X1-BaaS Scraping API
Stealth scraping API for AI agents. Clean Markdown from any URL. x402 crypto payments.
我该使用它吗
质量与安全性
基于对工具定义和协议合规性的自动分析。
上下文开销
这是每次将服务器的工具加载到模型上下文窗口时所消耗的大致 token 数。数值越高,可用于其他任务的注意力就越少。
安装
一键安装
将以下内容添加到你的 `claude_desktop_config.json` 文件中:
{
"mcpServers": {
"x1-baas": {
"url": "https://api.tazpal.com/mcp"
}
}
}远程端点
https://api.tazpal.com/mcpstreamable-http它能做什么
工具清单
工具(3)
🟢scrape(url, output, wait_for_selector, timeout_ms, block_media, ...)
Scrape a URL and return content in your preferred format. Supported output formats: - markdown (default): Clean LLM-ready Markdown text - screenshot: PNG/JPEG image of the page - pdf: PDF document of the page - csv: Table data extracted as CSV - html: Sanitized HTML with scripts/ads removed This tool handles: - JavaScript rendering (SPA, dynamic content) - Anti-bot bypass (Cloudflare Turnstile, Datadome) - DOM cleaning (strips scripts, nav, footer, ads) - HTML-to-Markdown conversion (Mozilla Readability engine) - Automatic retry with escalating wait strategies - Domain cooldown to avoid rate-limiting - Response caching (5 min TTL) Args: url: The URL to scrape (must start with http:// or https://) output: Output format: "markdown" (default), "screenshot", "pdf", "csv", "html" wait_for_selector: Optional CSS selector to wait for before extraction (e.g., ".article-content") timeout_ms: Navigation timeout in milliseconds (default: 20000, max: 120000) block_media: Block images/fonts/video for faster loading (default: true) wait_strategy: Wait strategy: "default", "spa", "heavy", "cloudflare" (auto-detected if omitted) retry: Enable automatic retry on failure (default: true) bypass_cache: Skip cache, force fresh scrape (default: false) javascript: Custom JavaScript to execute after page load (e.g., "window.scrollTo(0, 1000)") Returns: Content in the requested format, or an error message.
输入模式
{
"type": "object",
"properties": {
"url": {
"title": "Url",
"type": "string"
},
"output": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Output"
},
"wait_for_selector": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Wait For Selector"
},
"timeout_ms": {
"default": 20000,
"title": "Timeout Ms",
"type": "integer"
},
"block_media": {
"default": true,
"title": "Block Media",
"type": "boolean"
},
"wait_strategy": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Wait Strategy"
},
"retry": {
"default": true,
"title": "Retry",
"type": "boolean"
},
"bypass_cache": {
"default": false,
"title": "Bypass Cache",
"type": "boolean"
},
"javascript": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Javascript"
}
},
"required": [
"url"
],
"title": "scrapeArguments"
}输出模式
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "scrapeOutput"
}🟢get_pricing
Check current pricing and payment requirements for the BaaS scrape engine. Returns pricing info, supported networks, and payment instructions. No authentication required. Returns: Current pricing details and payment instructions.
输入模式
{
"type": "object",
"properties": {},
"title": "get_pricingArguments"
}输出模式
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "get_pricingOutput"
}🟢server_status
Check the BaaS engine health and operational status. Returns browser status, uptime, x402 payment mode, and context count. Returns: Server health and diagnostic information.
输入模式
{
"type": "object",
"properties": {},
"title": "server_statusArguments"
}输出模式
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "server_statusOutput"
}社区
证据