MCP server for vLLM - expose vLLM capabilities to AI assistants
Wellknown found it in public sources; nobody has proven control of it yet. Claiming takes one click if the repository is under your GitHub account, or a small file on your domain otherwise. Verified owners get the badge, 15-minute checks, status alerts, edits that outrank crawled data, and a ranking boost.
Agents can do it too: POST https://wellknown.network/api/v1/claims with {"agent":"vllm-mcp-server","method":"well_known_file"} — machine-readable steps at claim.json, guide at /docs/claim.
Everything here was measured by our prober or read from a registry. Nothing is self-reported.
Attributed to the source that supplied each field. Treated as claims, not facts.
# vLLM MCP Server [](https://www.python.org/downloads/) [](https://opensource.org/licenses/Apache-2.0) A [Model Context Protocol (MCP)](https://modelcontextprotocol.io/) server that exposes vLLM capabilities to AI assistants like Claude, Cursor, and other MCP-compatible clients. ## Features - 🚀 **Chat & Completion**: Send chat messages and text completions to vLLM - 📋 **Model Management**: List and inspect available models - 📊 **Server Monitoring**: Check server health and performance metrics - 🐳 **Platform-Aware Container Control**: Supports both Podman and Docker. Automatically detects your platform (Linux/macOS/Windows) and GPU availability, selecting the appropriate container image and optimal settings (e.g., `max_model_len`) - 📈 **Benchmarking**: Run GuideLLM benchmarks (optional) - 💬 **Pre-defined Prompts**: Use curated system prompts for common tasks ## Demo ### Start vLLM Server Use the `start_vllm` tool to launch a vLLM container with automatic platform detection:  ### Chat with vLLM Send chat messages using the `vllm_chat` tool:  ### Stop vLLM Server Clean up with the `stop_vllm` tool:  ## Installation ### Using uvx (Recommended) ```bash uvx vllm-mcp-server ``` ### Using pip ```bash pip install vllm-mcp-server ``` ### From Source ```bash git clone https://github.com/micytao/vllm-mcp-server.git cd vllm-mcp-server pip install -e . ``` ## Quick Start ### 1. Start a vLLM Server You can either start a vLLM server manually or let the MCP server manage it via Docker. …
Mapped onto the structured taxonomy from declared text and observed tool names. Confidence shown for derived entries.
Every source is kept verbatim. Field changes are logged as events.