Siteiz

Check which AI crawlers a site allows and see pages the way AI crawlers do. Free, read-only.

我该使用它吗

质量与安全性

A
描述质量
100%
模式完整度
95%
命名质量
90%
投毒风险
100%
权限匹配度
100%
协议合规性
100%

基于对工具定义和协议合规性的自动分析。

上下文开销

~1,197token 数(工具定义)
~2.2 KB典型响应大小
对注意力有中等影响(占 128k 上下文窗口的 0.94%)

这是每次将服务器的工具加载到模型上下文窗口时所消耗的大致 token 数。数值越高,可用于其他任务的注意力就越少。

安装

一键安装

将以下内容添加到你的 `claude_desktop_config.json` 文件中:

{
  "mcpServers": {
    "siteiz": {
      "url": "https://siteiz.com/mcp"
    }
  }
}

远程端点

https://siteiz.com/mcpstreamable-http

它能做什么

工具清单

工具(4)

🟢 只读🟡 写入🔴 删除⚪ 未知
🟢check_ai_crawlers(url)

Reads a website's robots.txt and reports, for 11 AI crawlers (GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, Claude-SearchBot, PerplexityBot, Google-Extended, Applebot-Extended, Meta-ExternalAgent, CCBot, Bytespider), whether each may reach the given page. Separates AI search crawlers (decide if the site appears in ChatGPT, Claude and Perplexity answers) from training crawlers. Also reports llms.txt and declared sitemaps.

输入模式

{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "Website or page address, e.g. example.com or https://example.com/pricing"
    }
  },
  "required": [
    "url"
  ],
  "additionalProperties": false
}

输出模式

{
  "type": "object",
  "properties": {
    "url": {
      "type": [
        "string",
        "null"
      ]
    },
    "robotsTxt": {
      "type": "string",
      "enum": [
        "ok",
        "missing",
        "unreachable",
        "invalid",
        "pasted"
      ]
    },
    "httpStatus": {
      "type": [
        "integer",
        "null"
      ]
    },
    "agents": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "token": {
            "type": "string"
          },
          "verdict": {
            "type": "string",
            "enum": [
              "allowed",
              "default",
              "blocked"
            ]
          },
          "operator": {
            "type": "string"
          },
          "role": {
            "type": "string",
            "enum": [
              "search",
              "user",
              "training",
              "control"
            ]
          }
        },
        "required": [
          "token",
          "verdict",
          "operator",
          "role"
        ]
      }
    },
    "llmsTxt": {
      "type": [
        "boolean",
        "null"
      ]
    },
    "sitemaps": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "contentSignals": {
      "type": "array",
      "items": {
        "type": "string"
      }
    }
  },
  "required": [
    "url",
    "robotsTxt",
    "agents"
  ]
}
🟢view_page_as_ai_crawler(url)

Fetches one page the way most AI crawlers do (plain HTTP, no JavaScript) and reports what they actually receive: title, meta description, canonical, H1, heading count, structured data types, visible word count, token reading cost, and a text preview. Use it to check whether content is visible to AI without JavaScript.

输入模式

{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "Website or page address, e.g. example.com or https://example.com/pricing"
    }
  },
  "required": [
    "url"
  ],
  "additionalProperties": false
}

输出模式

{
  "type": "object",
  "properties": {
    "finalUrl": {
      "type": "string"
    },
    "status": {
      "type": "integer"
    },
    "bytes": {
      "type": "integer"
    },
    "title": {
      "type": [
        "string",
        "null"
      ]
    },
    "metaDescription": {
      "type": [
        "string",
        "null"
      ]
    },
    "canonical": {
      "type": [
        "string",
        "null"
      ]
    },
    "h1": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "headingCount": {
      "type": "integer"
    },
    "structuredData": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "words": {
      "type": "integer"
    },
    "htmlTokens": {
      "type": "integer"
    },
    "textTokens": {
      "type": "integer"
    },
    "fullScan": {
      "type": "string"
    }
  },
  "required": [
    "finalUrl",
    "status",
    "words",
    "fullScan"
  ]
}
🟢explain_ai_crawler(name)

Explains what an AI crawler user agent does (operator, purpose, and what blocking it in robots.txt changes). Pass a name such as GPTBot or OAI-SearchBot, or omit it to list all crawlers Siteiz tracks.

输入模式

{
  "type": "object",
  "properties": {
    "name": {
      "type": "string",
      "description": "Crawler user agent, e.g. GPTBot. Omit to list all."
    }
  },
  "additionalProperties": false
}

输出模式

{
  "type": "object",
  "properties": {
    "crawlers": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "token": {
            "type": "string"
          },
          "operator": {
            "type": "string"
          },
          "product": {
            "type": "string"
          },
          "role": {
            "type": "string"
          },
          "roleLabel": {
            "type": "string"
          },
          "does": {
            "type": "string"
          }
        },
        "required": [
          "token",
          "operator",
          "role",
          "does"
        ]
      }
    }
  },
  "required": [
    "crawlers"
  ]
}
🟢get_ai_visibility_report(company)

Returns a published, dated Siteiz AI visibility report for a well-known company's homepage (score, grade, pillar scores, top issues). Omit the company to list all published reports.

输入模式

{
  "type": "object",
  "properties": {
    "company": {
      "type": "string",
      "description": "Company name, e.g. Notion. Omit to list all reports."
    }
  },
  "additionalProperties": false
}

输出模式

{
  "type": "object",
  "properties": {
    "reports": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "company": {
            "type": "string"
          },
          "score": {
            "type": "integer"
          },
          "grade": {
            "type": "string"
          },
          "scannedAt": {
            "type": "string"
          },
          "url": {
            "type": "string"
          }
        }
      }
    },
    "company": {
      "type": "string"
    },
    "url": {
      "type": "string"
    },
    "scannedAt": {
      "type": "string"
    },
    "score": {
      "type": "integer"
    },
    "grade": {
      "type": "string"
    },
    "pillars": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string"
          },
          "label": {
            "type": "string"
          },
          "score": {
            "type": "integer"
          },
          "notAssessed": {
            "type": "boolean"
          }
        }
      }
    },
    "report": {
      "type": "string"
    }
  },
  "description": "Either `reports` (a list, when no company matched) or the single report fields."
}

社区

评价此服务器

证据

最近观测

已验证未记录版本4 个工具