Stringer CleanExtract

Convert public HTML or a public URL into token-dense Markdown through one remote MCP tool.

Should I use this

Quality & Safety

A
Description quality
100%
Schema completeness
100%
Naming quality
80%
Poisoning risk
100%
Permission match
100%
Protocol compliance
100%

Based on automated analysis of tool definitions and protocol compliance.

Context Cost

~394Tokens (tool definitions)
~1.2 KBTypical response size
Minimal attention impact (0.31% of 128k context)

This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.

Install

One-Click Install

Add this to your `claude_desktop_config.json` file:

{
  "mcpServers": {
    "cleanextract": {
      "url": "https://extract.getstringer.app/mcp"
    }
  }
}

Remote endpoints

https://extract.getstringer.app/mcpstreamable-http

What it can do

Tool inventory

Tools (2)

🟢 Read-only🟡 Write🔴 Delete⚪ Unknown
🟢clean_extract(url_or_html, max_output_bytes, sections)

Extract a full page or selected sections from a public URL or raw HTML. A call is charged only when it returns usable content; the first 3 usable extraction calls are free, total, then USD 0.05. Use the free clean_extract_outline first. Uncharged failures name upstream_timeout, upstream_http_error, bot_challenge, js_shell, empty_extraction, or sections_not_found. Recoverable structured data names fallback_used. Every success reports extraction_quality and extraction_content_ratio_band.

Input Schema

{
  "type": "object",
  "properties": {
    "url_or_html": {
      "type": "string",
      "minLength": 1,
      "description": "A public HTTP(S) URL or a raw HTML string to convert into Markdown."
    },
    "max_output_bytes": {
      "type": "integer",
      "minimum": 1,
      "maximum": 1048576,
      "description": "Optional maximum UTF-8 byte length of the returned Markdown."
    },
    "sections": {
      "type": "array",
      "minItems": 1,
      "maxItems": 50,
      "items": {
        "type": "string",
        "pattern": "^s[1-9][0-9]*$"
      },
      "description": "Optional ids from clean_extract_outline. Return only these sections in document order."
    }
  },
  "required": [
    "url_or_html"
  ],
  "additionalProperties": false
}
🟢clean_extract_outline(url_or_html)

Free outline with section ids, headings, byte sizes, and a fingerprint, without page body text or payment. A later clean_extract call can select sections and is charged only when it returns usable content. Uncharged failures name upstream_timeout, upstream_http_error, bot_challenge, js_shell, empty_extraction, or sections_not_found. Recoverable structured data names fallback_used.

Input Schema

{
  "type": "object",
  "properties": {
    "url_or_html": {
      "type": "string",
      "minLength": 1,
      "description": "A public HTTP(S) URL or a raw HTML string to convert into Markdown."
    }
  },
  "required": [
    "url_or_html"
  ],
  "additionalProperties": false
}

Community

Rate this Server

Evidence

Recent observations

verifiedversion not recorded2 tools
verifiedversion not recorded1 tools