AI-powered document perception and analysis MCP server with intelligent provider selection
Wellknown found it in public sources; nobody has proven control of it yet. Claiming takes one click if the repository is under your GitHub account, or a small file on your domain otherwise. Verified owners get the badge, 15-minute checks, status alerts, edits that outrank crawled data, and a ranking boost.
Agents can do it too: POST https://wellknown.network/api/v1/claims with {"agent":"docsray-mcp","method":"well_known_file"} — machine-readable steps at claim.json, guide at /docs/claim.
Everything here was measured by our prober or read from a registry. Nothing is self-reported.
Attributed to the source that supplied each field. Treated as claims, not facts.
# 🔍 Docsray MCP Server [](https://pypi.org/project/docsray-mcp/) [](https://opensource.org/licenses/Apache-2.0) [](https://www.python.org/downloads/) [](https://github.com/anthropics/mcp) [](https://github.com/docsray/docsray-mcp) [](https://app.netlify.com/projects/docsray/deploys) **Docsray** is a powerful Model Context Protocol (MCP) server that gives AI assistants like Claude advanced document perception capabilities. Extract text, navigate pages, analyze structure, and understand any document with ease. **✅ Status: Published to PyPI and TestPyPI - Working in Cursor, Claude Desktop, and other MCP clients** ## ✨ Features ### 🎯 Five Powerful Tools 1. **`docsray_peek`** - Quick document overview with format detection and provider capabilities 2. **`docsray_map`** - Generate comprehensive document structure maps with caching 3. **`docsray_xray`** - AI-powered deep analysis extracting entities, relationships, and insights 4. **`docsray_extract`** - Extract content in multiple formats (markdown, text, JSON, tables) 5. **`docsray_seek`** - Navigate to specific pages, sections, or search for content ### 🔌 Multi-Provider Architecture - **PyMuPDF4LLM** - Lightning-fast PDF processing (✅ Implemented) - Fast markdown extraction - Basic table detection - Multi-page support - Always enabled as fallback - **LlamaParse** - Deep document understanding with LLMs (✅ Implemented) - AI-powered entity extraction - Custom analysis instructions - Comprehensive caching in .docsray directories - Rich format preservation (markdown, i…
Mapped onto the structured taxonomy from declared text and observed tool names. Confidence shown for derived entries.
Every source is kept verbatim. Field changes are logged as events.