Stringer CleanExtract
Convert public HTML or a public URL into token-dense Markdown through one remote MCP tool.
Should I use this
Quality & Safety
Based on automated analysis of tool definitions and protocol compliance.
Context Cost
This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.
Install
One-Click Install
Add this to your `claude_desktop_config.json` file:
{
"mcpServers": {
"cleanextract": {
"url": "https://extract.getstringer.app/mcp"
}
}
}Remote endpoints
https://extract.getstringer.app/mcpstreamable-httpWhat it can do
Tool inventory
Tools (2)
🟢clean_extract(url_or_html, max_output_bytes, sections)
Extract a full page or selected sections from a public URL or raw HTML. A call is charged only when it returns usable content; the first 3 usable extraction calls are free, total, then USD 0.05. Use the free clean_extract_outline first. Uncharged failures name upstream_timeout, upstream_http_error, bot_challenge, js_shell, empty_extraction, or sections_not_found. Recoverable structured data names fallback_used. Every success reports extraction_quality and extraction_content_ratio_band.
Input Schema
{
"type": "object",
"properties": {
"url_or_html": {
"type": "string",
"minLength": 1,
"description": "A public HTTP(S) URL or a raw HTML string to convert into Markdown."
},
"max_output_bytes": {
"type": "integer",
"minimum": 1,
"maximum": 1048576,
"description": "Optional maximum UTF-8 byte length of the returned Markdown."
},
"sections": {
"type": "array",
"minItems": 1,
"maxItems": 50,
"items": {
"type": "string",
"pattern": "^s[1-9][0-9]*$"
},
"description": "Optional ids from clean_extract_outline. Return only these sections in document order."
}
},
"required": [
"url_or_html"
],
"additionalProperties": false
}🟢clean_extract_outline(url_or_html)
Free outline with section ids, headings, byte sizes, and a fingerprint, without page body text or payment. A later clean_extract call can select sections and is charged only when it returns usable content. Uncharged failures name upstream_timeout, upstream_http_error, bot_challenge, js_shell, empty_extraction, or sections_not_found. Recoverable structured data names fallback_used.
Input Schema
{
"type": "object",
"properties": {
"url_or_html": {
"type": "string",
"minLength": 1,
"description": "A public HTTP(S) URL or a raw HTML string to convert into Markdown."
}
},
"required": [
"url_or_html"
],
"additionalProperties": false
}Community
Evidence