PDF to Markdown MCP服务器
Wellknown found it in public sources; nobody has proven control of it yet. Claiming takes one click if the repository is under your GitHub account, or a small file on your domain otherwise. Verified owners get the badge, 15-minute checks, status alerts, edits that outrank crawled data, and a ranking boost.
Agents can do it too: POST https://wellknown.network/api/v1/claims with {"agent":"iflow-mcp-pdf2md","method":"well_known_file"} — machine-readable steps at claim.json, guide at /docs/claim.
Everything here was measured by our prober or read from a registry. Nothing is self-reported.
Attributed to the source that supplied each field. Treated as claims, not facts.
# MCP-PDF2MD [](https://smithery.ai/server/@FutureUnreal/mcp-pdf2md) [English](#pdf2md-service) | [中文](README_CN.md) # MCP-PDF2MD Service An MCP-based high-performance PDF to Markdown conversion service powered by MinerU API, supporting batch processing for local files and URL links with structured output. ## Key Features - Format Conversion: Convert PDF files to structured Markdown format. - Multi-source Support: Process both local PDF files and URL links. - Intelligent Processing: Automatically select the best processing method. - Batch Processing: Support multi-file batch conversion for efficient handling of large volumes of PDF files. - MCP Integration: Seamless integration with LLM clients like Claude Desktop. - Structure Preservation: Maintain the original document structure, including headings, paragraphs, lists, etc. - Smart Layout: Output text in human-readable order, suitable for single-column, multi-column, and complex layouts. - Formula Conversion: Automatically recognize and convert formulas in the document to LaTeX format. - Table Extraction: Automatically recognize and convert tables in the document to structured format. - Cleanup Optimization: Remove headers, footers, footnotes, page numbers, etc., to ensure semantic coherence. - High-Quality Extraction: High-quality extraction of text, images, and layout information from PDF documents. ## System Requirements - Software: Python 3.10+ ## Quick Start 1. Clone the repository and enter the directory: ```bash git clone https://github.com/FutureUnreal/mcp-pdf2md.git cd mcp-pdf2md ``` 2. Create a virtual environment and install dependencies: **Linux/macOS**: ```bash uv venv source .venv/bin/activate uv pip install -e . ``` **Windows**: ```bash uv venv .venv\Scripts\activate uv pip install -e . ``` 3. Configure environment variables: Create a `.env` file in the project …
Mapped onto the structured taxonomy from declared text and observed tool names. Confidence shown for derived entries.
Every source is kept verbatim. Field changes are logged as events.