Siteiz
Check which AI crawlers a site allows and see pages the way AI crawlers do. Free, read-only.
我该使用它吗
质量与安全性
基于对工具定义和协议合规性的自动分析。
上下文开销
这是每次将服务器的工具加载到模型上下文窗口时所消耗的大致 token 数。数值越高,可用于其他任务的注意力就越少。
安装
一键安装
将以下内容添加到你的 `claude_desktop_config.json` 文件中:
{
"mcpServers": {
"siteiz": {
"url": "https://siteiz.com/mcp"
}
}
}远程端点
https://siteiz.com/mcpstreamable-http它能做什么
工具清单
工具(4)
🟢check_ai_crawlers(url)
Reads a website's robots.txt and reports, for 11 AI crawlers (GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, Claude-SearchBot, PerplexityBot, Google-Extended, Applebot-Extended, Meta-ExternalAgent, CCBot, Bytespider), whether each may reach the given page. Separates AI search crawlers (decide if the site appears in ChatGPT, Claude and Perplexity answers) from training crawlers. Also reports llms.txt and declared sitemaps.
输入模式
{
"type": "object",
"properties": {
"url": {
"type": "string",
"description": "Website or page address, e.g. example.com or https://example.com/pricing"
}
},
"required": [
"url"
],
"additionalProperties": false
}输出模式
{
"type": "object",
"properties": {
"url": {
"type": [
"string",
"null"
]
},
"robotsTxt": {
"type": "string",
"enum": [
"ok",
"missing",
"unreachable",
"invalid",
"pasted"
]
},
"httpStatus": {
"type": [
"integer",
"null"
]
},
"agents": {
"type": "array",
"items": {
"type": "object",
"properties": {
"token": {
"type": "string"
},
"verdict": {
"type": "string",
"enum": [
"allowed",
"default",
"blocked"
]
},
"operator": {
"type": "string"
},
"role": {
"type": "string",
"enum": [
"search",
"user",
"training",
"control"
]
}
},
"required": [
"token",
"verdict",
"operator",
"role"
]
}
},
"llmsTxt": {
"type": [
"boolean",
"null"
]
},
"sitemaps": {
"type": "array",
"items": {
"type": "string"
}
},
"contentSignals": {
"type": "array",
"items": {
"type": "string"
}
}
},
"required": [
"url",
"robotsTxt",
"agents"
]
}🟢view_page_as_ai_crawler(url)
Fetches one page the way most AI crawlers do (plain HTTP, no JavaScript) and reports what they actually receive: title, meta description, canonical, H1, heading count, structured data types, visible word count, token reading cost, and a text preview. Use it to check whether content is visible to AI without JavaScript.
输入模式
{
"type": "object",
"properties": {
"url": {
"type": "string",
"description": "Website or page address, e.g. example.com or https://example.com/pricing"
}
},
"required": [
"url"
],
"additionalProperties": false
}输出模式
{
"type": "object",
"properties": {
"finalUrl": {
"type": "string"
},
"status": {
"type": "integer"
},
"bytes": {
"type": "integer"
},
"title": {
"type": [
"string",
"null"
]
},
"metaDescription": {
"type": [
"string",
"null"
]
},
"canonical": {
"type": [
"string",
"null"
]
},
"h1": {
"type": "array",
"items": {
"type": "string"
}
},
"headingCount": {
"type": "integer"
},
"structuredData": {
"type": "array",
"items": {
"type": "string"
}
},
"words": {
"type": "integer"
},
"htmlTokens": {
"type": "integer"
},
"textTokens": {
"type": "integer"
},
"fullScan": {
"type": "string"
}
},
"required": [
"finalUrl",
"status",
"words",
"fullScan"
]
}🟢explain_ai_crawler(name)
Explains what an AI crawler user agent does (operator, purpose, and what blocking it in robots.txt changes). Pass a name such as GPTBot or OAI-SearchBot, or omit it to list all crawlers Siteiz tracks.
输入模式
{
"type": "object",
"properties": {
"name": {
"type": "string",
"description": "Crawler user agent, e.g. GPTBot. Omit to list all."
}
},
"additionalProperties": false
}输出模式
{
"type": "object",
"properties": {
"crawlers": {
"type": "array",
"items": {
"type": "object",
"properties": {
"token": {
"type": "string"
},
"operator": {
"type": "string"
},
"product": {
"type": "string"
},
"role": {
"type": "string"
},
"roleLabel": {
"type": "string"
},
"does": {
"type": "string"
}
},
"required": [
"token",
"operator",
"role",
"does"
]
}
}
},
"required": [
"crawlers"
]
}🟢get_ai_visibility_report(company)
Returns a published, dated Siteiz AI visibility report for a well-known company's homepage (score, grade, pillar scores, top issues). Omit the company to list all published reports.
输入模式
{
"type": "object",
"properties": {
"company": {
"type": "string",
"description": "Company name, e.g. Notion. Omit to list all reports."
}
},
"additionalProperties": false
}输出模式
{
"type": "object",
"properties": {
"reports": {
"type": "array",
"items": {
"type": "object",
"properties": {
"company": {
"type": "string"
},
"score": {
"type": "integer"
},
"grade": {
"type": "string"
},
"scannedAt": {
"type": "string"
},
"url": {
"type": "string"
}
}
}
},
"company": {
"type": "string"
},
"url": {
"type": "string"
},
"scannedAt": {
"type": "string"
},
"score": {
"type": "integer"
},
"grade": {
"type": "string"
},
"pillars": {
"type": "array",
"items": {
"type": "object",
"properties": {
"id": {
"type": "string"
},
"label": {
"type": "string"
},
"score": {
"type": "integer"
},
"notAssessed": {
"type": "boolean"
}
}
}
},
"report": {
"type": "string"
}
},
"description": "Either `reports` (a list, when no company matched) or the single report fields."
}社区
证据