docforge

12 paid document tools: PDF to markdown, OCR, tables, invoices, Word/Excel/HTML, merge/split. x402

我該用這個嗎

品質與安全性

A
說明品質
100%
結構描述完整度
92%
命名品質
88%
汙染風險
100%
權限相符程度
100%
協定合規性
100%

根據工具定義與協定合規性的自動化分析。

上下文成本

~756Token(工具定義)
~1.1 KB典型回應大小
中等的注意力影響(128k 上下文的 0.59%)

這是每次將伺服器的工具載入模型上下文時所消耗的約略 token 數量。數量越高,可用於其他工作的注意力就越少。

安裝

一鍵安裝

將以下內容加入你的 `claude_desktop_config.json` 檔案:

{
  "mcpServers": {
    "docforge": {
      "url": "https://docforged.mcpize.run/mcp"
    }
  }
}

遠端端點

https://docforged.mcpize.run/mcpstreamable-http

它能做什麼

工具清單

工具(5)

🟢 唯讀🟡 寫入🔴 刪除⚪ 未知
🟢pdf_to_markdown(file_url, file_base64)

Extract the text of a PDF and convert it to clean markdown. Detects headings by font size and preserves lists and paragraphs. Input: a text-based PDF via file_url or file_base64. For scanned PDFs use ocr_image on page images instead.

輸入結構描述

{
  "type": "object",
  "properties": {
    "file_url": {
      "type": "string",
      "format": "uri",
      "description": "Public http(s) URL of the file"
    },
    "file_base64": {
      "type": "string",
      "description": "Base64-encoded file contents (data-URI prefix allowed)"
    }
  },
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
⚪ocr_image(file_url, file_base64, language)

Run optical character recognition on an image (png, jpg, webp, bmp) and return the recognized text with a confidence score. Supports 100+ languages via the language parameter (ISO 639-2 codes like 'eng', 'deu', 'fra', 'spa').

輸入結構描述

{
  "type": "object",
  "properties": {
    "file_url": {
      "type": "string",
      "format": "uri",
      "description": "Public http(s) URL of the file"
    },
    "file_base64": {
      "type": "string",
      "description": "Base64-encoded file contents (data-URI prefix allowed)"
    },
    "language": {
      "type": "string",
      "description": "Tesseract language code, default 'eng'"
    }
  },
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🔴extract_tables(file_url, file_base64)

Detect and reconstruct tables from a text-based PDF. Returns each table as structured rows plus ready-to-use markdown and CSV renderings. Works best on PDFs with clear columnar layout (invoices, reports, statements).

輸入結構描述

{
  "type": "object",
  "properties": {
    "file_url": {
      "type": "string",
      "format": "uri",
      "description": "Public http(s) URL of the file"
    },
    "file_base64": {
      "type": "string",
      "description": "Base64-encoded file contents (data-URI prefix allowed)"
    }
  },
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
⚪render_pdf(content, format, title)

Render markdown (or simple HTML) into a clean, printable A4 PDF. Supports headings, paragraphs, bullet and numbered lists, blockquotes, code blocks, horizontal rules, and inline bold/italic/code. Returns the PDF as base64 plus page count.

輸入結構描述

{
  "type": "object",
  "properties": {
    "content": {
      "type": "string",
      "minLength": 1,
      "description": "The markdown or HTML source to render"
    },
    "format": {
      "type": "string",
      "enum": [
        "markdown",
        "html"
      ],
      "description": "Input format, default markdown"
    },
    "title": {
      "type": "string",
      "description": "PDF document title metadata"
    }
  },
  "required": [
    "content"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢parse_invoice(file_url, file_base64, is_image)

Extract structured data from an invoice or receipt: vendor, invoice number, dates, currency, subtotal, tax, total, and line items. Accepts a text-based PDF, or an image when is_image is true (OCR is applied first). Returns JSON.

輸入結構描述

{
  "type": "object",
  "properties": {
    "file_url": {
      "type": "string",
      "format": "uri",
      "description": "Public http(s) URL of the file"
    },
    "file_base64": {
      "type": "string",
      "description": "Base64-encoded file contents (data-URI prefix allowed)"
    },
    "is_image": {
      "type": "boolean",
      "description": "Set true when the file is a photo/scan image rather than a PDF"
    }
  },
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}

社群

為此伺服器評分

證據

近期觀測

已驗證未記錄版本5 個工具
已驗證未記錄版本5 個工具
已驗證未記錄版本5 個工具