This is a local rag-mcp solution with chromadb using langchain and docling
Wellknown found it in public sources; nobody has proven control of it yet. Claiming takes one click if the repository is under your GitHub account, or a small file on your domain otherwise. Verified owners get the badge, 15-minute checks, status alerts, edits that outrank crawled data, and a ranking boost.
Agents can do it too: POST https://wellknown.network/api/v1/claims with {"agent":"rag-mcp-2","method":"well_known_file"} — machine-readable steps at claim.json, guide at /docs/claim.
Everything here was measured by our prober or read from a registry. Nothing is self-reported.
Attributed to the source that supplied each field. Treated as claims, not facts.
# RAG MCP: Document Processing Server A Retrieval-Augmented Generation (RAG) server built on the Model Context Protocol (MCP) for intelligent document processing and question answering. ## Overview RAG MCP is a tool that allows you to index various document formats and perform semantic searches against them. It uses advanced embedding techniques and vector databases to make your documents searchable through natural language queries. ## Features - **Document Indexing**: Support for various document formats (PDF, DOCX, XLSX, PPTX, Markdown, AsciiDoc, HTML, XHTML, CSV) - **Semantic Search**: Query your documents using natural language - **Flexible Embedding Models**: Choose between HuggingFace BGE (default) or Ollama embeddings. - **High Performance**: Optimized for various hardware configurations with automatic device selection (CUDA, MPS, CPU) for HuggingFace embeddings. - **Persistent Storage**: Vector embeddings are stored locally for future use ## Requirements - Python 3.11+ - Environment with access to your documents - (Optional) Ollama installed and running if using Ollama embeddings. ## Installation ### 1. Install UV First, you need to install UV, a Python package installer and resolver: #### On macOS/Linux: ```bash curl -sSf https://astral.sh/uv/install.sh | sh ``` #### On Windows: ```bash powershell -c "irm https://astral.sh/uv/install.ps1 | iex" ``` ### 2. Run RAG MCP Once UV is installed, you can run RAG MCP directly using: ```bash uvx rag-mcp ``` This will start the MCP server and make it available for document processing. **Environment Variables:** You can configure RAG MCP using environment variables: - `PERSIST_DIRECTORY` (Required): Path to the directory where the vector database will be stored (e.g., `/path/to/your/persist/directory`). A `chromadb` subfolder will be created here. - `USE_OLLAMA_EMBEDDING` (Optional): Set to `True` to use Ollama embeddings instead of the default HuggingFace BGE embeddings. Requires Ollama to be r…
Mapped onto the structured taxonomy from declared text and observed tool names. Confidence shown for derived entries.
Every source is kept verbatim. Field changes are logged as events.