Siteiz

Check which AI crawlers a site allows and see pages the way AI crawlers do. Free, read-only.

¿Debería usar esto?

Calidad y seguridad

A
Calidad de la descripción
100%
Integridad del esquema
95%
Calidad de los nombres
90%
Riesgo de envenenamiento
100%
Coincidencia de permisos
100%
Cumplimiento del protocolo
100%

Basado en el análisis automatizado de las definiciones de herramientas y el cumplimiento del protocolo.

Costo de contexto

~1,197Tokens (definiciones de herramientas)
~2.2 KBTamaño de respuesta típico
Impacto moderado en la atención (0.94% del contexto de 128k)

Este es el número aproximado de tokens que se consumen cada vez que las herramientas del servidor se cargan en el contexto de un modelo. Los recuentos más altos reducen la atención disponible para otras tareas.

Instalar

Instalación con un clic

Agrega esto a tu archivo `claude_desktop_config.json`:

{
  "mcpServers": {
    "siteiz": {
      "url": "https://siteiz.com/mcp"
    }
  }
}

Puntos de conexión remotos

https://siteiz.com/mcpstreamable-http

Qué puede hacer

Inventario de herramientas

Herramientas (4)

🟢 Solo lectura🟡 Escritura🔴 Eliminación⚪ Desconocido
🟢check_ai_crawlers(url)

Reads a website's robots.txt and reports, for 11 AI crawlers (GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, Claude-SearchBot, PerplexityBot, Google-Extended, Applebot-Extended, Meta-ExternalAgent, CCBot, Bytespider), whether each may reach the given page. Separates AI search crawlers (decide if the site appears in ChatGPT, Claude and Perplexity answers) from training crawlers. Also reports llms.txt and declared sitemaps.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "Website or page address, e.g. example.com or https://example.com/pricing"
    }
  },
  "required": [
    "url"
  ],
  "additionalProperties": false
}

Esquema de salida

{
  "type": "object",
  "properties": {
    "url": {
      "type": [
        "string",
        "null"
      ]
    },
    "robotsTxt": {
      "type": "string",
      "enum": [
        "ok",
        "missing",
        "unreachable",
        "invalid",
        "pasted"
      ]
    },
    "httpStatus": {
      "type": [
        "integer",
        "null"
      ]
    },
    "agents": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "token": {
            "type": "string"
          },
          "verdict": {
            "type": "string",
            "enum": [
              "allowed",
              "default",
              "blocked"
            ]
          },
          "operator": {
            "type": "string"
          },
          "role": {
            "type": "string",
            "enum": [
              "search",
              "user",
              "training",
              "control"
            ]
          }
        },
        "required": [
          "token",
          "verdict",
          "operator",
          "role"
        ]
      }
    },
    "llmsTxt": {
      "type": [
        "boolean",
        "null"
      ]
    },
    "sitemaps": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "contentSignals": {
      "type": "array",
      "items": {
        "type": "string"
      }
    }
  },
  "required": [
    "url",
    "robotsTxt",
    "agents"
  ]
}
🟢view_page_as_ai_crawler(url)

Fetches one page the way most AI crawlers do (plain HTTP, no JavaScript) and reports what they actually receive: title, meta description, canonical, H1, heading count, structured data types, visible word count, token reading cost, and a text preview. Use it to check whether content is visible to AI without JavaScript.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "Website or page address, e.g. example.com or https://example.com/pricing"
    }
  },
  "required": [
    "url"
  ],
  "additionalProperties": false
}

Esquema de salida

{
  "type": "object",
  "properties": {
    "finalUrl": {
      "type": "string"
    },
    "status": {
      "type": "integer"
    },
    "bytes": {
      "type": "integer"
    },
    "title": {
      "type": [
        "string",
        "null"
      ]
    },
    "metaDescription": {
      "type": [
        "string",
        "null"
      ]
    },
    "canonical": {
      "type": [
        "string",
        "null"
      ]
    },
    "h1": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "headingCount": {
      "type": "integer"
    },
    "structuredData": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "words": {
      "type": "integer"
    },
    "htmlTokens": {
      "type": "integer"
    },
    "textTokens": {
      "type": "integer"
    },
    "fullScan": {
      "type": "string"
    }
  },
  "required": [
    "finalUrl",
    "status",
    "words",
    "fullScan"
  ]
}
🟢explain_ai_crawler(name)

Explains what an AI crawler user agent does (operator, purpose, and what blocking it in robots.txt changes). Pass a name such as GPTBot or OAI-SearchBot, or omit it to list all crawlers Siteiz tracks.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "name": {
      "type": "string",
      "description": "Crawler user agent, e.g. GPTBot. Omit to list all."
    }
  },
  "additionalProperties": false
}

Esquema de salida

{
  "type": "object",
  "properties": {
    "crawlers": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "token": {
            "type": "string"
          },
          "operator": {
            "type": "string"
          },
          "product": {
            "type": "string"
          },
          "role": {
            "type": "string"
          },
          "roleLabel": {
            "type": "string"
          },
          "does": {
            "type": "string"
          }
        },
        "required": [
          "token",
          "operator",
          "role",
          "does"
        ]
      }
    }
  },
  "required": [
    "crawlers"
  ]
}
🟢get_ai_visibility_report(company)

Returns a published, dated Siteiz AI visibility report for a well-known company's homepage (score, grade, pillar scores, top issues). Omit the company to list all published reports.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "company": {
      "type": "string",
      "description": "Company name, e.g. Notion. Omit to list all reports."
    }
  },
  "additionalProperties": false
}

Esquema de salida

{
  "type": "object",
  "properties": {
    "reports": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "company": {
            "type": "string"
          },
          "score": {
            "type": "integer"
          },
          "grade": {
            "type": "string"
          },
          "scannedAt": {
            "type": "string"
          },
          "url": {
            "type": "string"
          }
        }
      }
    },
    "company": {
      "type": "string"
    },
    "url": {
      "type": "string"
    },
    "scannedAt": {
      "type": "string"
    },
    "score": {
      "type": "integer"
    },
    "grade": {
      "type": "string"
    },
    "pillars": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string"
          },
          "label": {
            "type": "string"
          },
          "score": {
            "type": "integer"
          },
          "notAssessed": {
            "type": "boolean"
          }
        }
      }
    },
    "report": {
      "type": "string"
    }
  },
  "description": "Either `reports` (a list, when no company matched) or the single report fields."
}

Comunidad

Califica este servidor

Evidencia

Observaciones recientes

verificadoversión no registrada4 herramientas