检索 24好玩 H5 互动营销平台的活动模板库、客户案例库与帮助知识库。匿名只读,无需 API Key。
Evaluate model or agent outputs.
检索 24好玩 H5 互动营销平台的活动模板库、客户案例库与帮助知识库。匿名只读,无需 API Key。
45 specialized judges that evaluate AI-generated code for security, cost, and quality.
Deterministic regression testing, release bisection and plugin compatibility matrices for Claude Code
Dependency intelligence for AI coding agents. Live CVE scanning, dependency health, upgrade planning, ecosystem news, decision memory — zero config, privacy-first. 9 tools standalone, 14 total with desktop app.
Persistent Project Context for Claude. IANA-registered .faf format, 34 MCP tools, 594 tests.
The adaptive AI harness for React Native — project-aware SDLC intelligence for any agent
Local-first repo-memory MCP for coding agents. Best for relationship-heavy edits: symbols, callers, dependencies, diff review, and git-pinned decisions before code changes. Start with a no-write proof receipt; use grep/ripgrep for exact strings. MIT; bund
Search and retrieve US court opinions, federal dockets, judge records, citation networks, and oral arguments from CourtListener's 9M+ opinion corpus via MCP. STDIO or Streamable HTTP.
MCP server for EdgeGate — set up edge-AI regression gates from Claude Code, Cursor, or Claude Desktop.
Author verifiable eval records through a draft → review → revise → submit loop with server-enforced graders; compile to JSONL/CSV/Inspect/lm-eval via MCP. STDIO or Streamable HTTP.
Investigate fraud from Claude, Cursor or any MCP client. Explain any verdict with the evidence behind it, pivot from one signup to every account sharing its device, IP or inbox, check an entity against a cross-operator abuse network, and work the review q
Eval-backed tool discovery for AI agents on the auxiliar.ai web-access gateway. recommend_tools picks the best search, scraping, browser-automation or voice provider for a job from measured benchmarks (quality, latency, cost, error rate), returning routes
Model Context Protocol server for ZeroBounce email validation, verification, and intelligence. Provides comprehensive email validation, AI scoring, email finder, domain search, and bulk processing capabilities for AI coding assistants.
MCP server for LLM/VLM model selection — compare 300+ models with real-time benchmarks, pricing, and personalized recommendations. No API key required.
AINative ZeroDB Memory MCP Server - 18 tools for agent memory: 9 memory + 5 write-back (Slack, Gmail, Calendar, GitHub, Notion) + 4 plan artifacts (persistent plans/PRDs across sessions with diff history). Auto-context middleware. Production-ready for Cla
Multi-model consensus: 2-6 frontier LLMs answer, an independent judge synthesises one answer.
MCP server that exposes the Langfuse REST API as tools — query traces, observations, sessions, scores, prompts, datasets, and metrics from any MCP client.
Stop shipping agents on vibes. Score every agent output for quality, safety, and cost.
Guarded CLI and MCP client for JustHandled's x402 utility gateway
Autonomous MCP server for Omniology — real-USDC AI agent skill contests on Solana mainnet. A live benchmark against real agents, every 88 seconds, 24/7.