Stringer CleanExtract
Convert public HTML or a public URL into token-dense Markdown through one remote MCP tool.
¿Debería usar esto?
Calidad y seguridad
Basado en el análisis automatizado de las definiciones de herramientas y el cumplimiento del protocolo.
Costo de contexto
Este es el número aproximado de tokens que se consumen cada vez que las herramientas del servidor se cargan en el contexto de un modelo. Los recuentos más altos reducen la atención disponible para otras tareas.
Instalar
Instalación con un clic
Agrega esto a tu archivo `claude_desktop_config.json`:
{
"mcpServers": {
"cleanextract": {
"url": "https://extract.getstringer.app/mcp"
}
}
}Puntos de conexión remotos
https://extract.getstringer.app/mcpstreamable-httpQué puede hacer
Inventario de herramientas
Herramientas (2)
🟢clean_extract(url_or_html, max_output_bytes, sections)
Extract a full page or selected sections from a public URL or raw HTML. A call is charged only when it returns usable content; the first 3 usable extraction calls are free, total, then USD 0.05. Use the free clean_extract_outline first. Uncharged failures name upstream_timeout, upstream_http_error, bot_challenge, js_shell, empty_extraction, or sections_not_found. Recoverable structured data names fallback_used. Every success reports extraction_quality and extraction_content_ratio_band.
Esquema de entrada
{
"type": "object",
"properties": {
"url_or_html": {
"type": "string",
"minLength": 1,
"description": "A public HTTP(S) URL or a raw HTML string to convert into Markdown."
},
"max_output_bytes": {
"type": "integer",
"minimum": 1,
"maximum": 1048576,
"description": "Optional maximum UTF-8 byte length of the returned Markdown."
},
"sections": {
"type": "array",
"minItems": 1,
"maxItems": 50,
"items": {
"type": "string",
"pattern": "^s[1-9][0-9]*$"
},
"description": "Optional ids from clean_extract_outline. Return only these sections in document order."
}
},
"required": [
"url_or_html"
],
"additionalProperties": false
}🟢clean_extract_outline(url_or_html)
Free outline with section ids, headings, byte sizes, and a fingerprint, without page body text or payment. A later clean_extract call can select sections and is charged only when it returns usable content. Uncharged failures name upstream_timeout, upstream_http_error, bot_challenge, js_shell, empty_extraction, or sections_not_found. Recoverable structured data names fallback_used.
Esquema de entrada
{
"type": "object",
"properties": {
"url_or_html": {
"type": "string",
"minLength": 1,
"description": "A public HTTP(S) URL or a raw HTML string to convert into Markdown."
}
},
"required": [
"url_or_html"
],
"additionalProperties": false
}Comunidad
Evidencia