Every change Wellknown observed on this MCP server, newest first, with what it was before and what it became. Tool-surface changes carry the definition diff. Nothing here is edited after the fact.
Changed the definition of "compare_capabilities", "evaluate_capability", "report_outcome" and 2 more
⟨19 unchanged words⟩ ,"type":"number"},"capabilities":{"description":"Two to five capabilities to compare, each as an id, ak:// URI or slug.","items":{"description":"Capability id (ak: ⟨18 unchanged words⟩ ,"type":"array"},"constraints":{"description":"Limits you set for the answer; a capability outside them is reported as blocked, with the reason code.","properties":{"max_cost_usd":{"description":"The highest price in USD you accept; a capability priced above it is blocked (over_max_cost).","minimum":0,"type":"number"},"minimum_confidence":{"description":"The lowest confidence in the measured saving you accept, from 0.5 to 0.999 (default 0.9); below it the answer is blocked (below_minimum_confidence).","maximum":0.999,"minimum":0.5,"type": ⟨84 unchanged words⟩
⟨39 unchanged words⟩ human approval.","properties":{"absolute_purchase_limit_usd":{"description":"The highest price in USD for one purchase; a price above it is outside the policy.","type":"number"},"auto_purchase":{"description":"When your operator allows a purchase without asking first: enabled (true or false), max_price_usd, and minimum_expected_roi (expected net saving divided by the price). The answer says whether the policy would allow it; these tools never charge.","type":"object"},"human_approval_required":{"description":"above_usd: a price above this amount in USD needs your operator's approval even where auto_purchase is enabled.","type":"object"},"monthly_budget_usd":{"description":"Your operator's budget for this month in USD.","type":"number"},"spent_this_month_usd":{"description":"What is already spent of that budget this month in USD (0 when left out).","type":"number"}},"type":"object" ⟨17 unchanged words⟩ ,"type":"string"},"constraints":{"description":"Limits you set for the answer; a capability outside them is reported as blocked, with the reason code.","properties":{"max_cost_usd":{"description":"The highest price in USD you accept; a capability priced above it is blocked (over_max_cost).","minimum":0,"type":"number"},"minimum_confidence":{"description":"The lowest confidence in the measured saving you accept, from 0.5 to 0.999 (default 0.9); below it the answer is blocked (below_minimum_confidence).","maximum":0.999,"minimum":0.5,"type": ⟨84 unchanged words⟩
{"additionalProperties":false,"properties":{"baseline_cost_estimate_usd":{"description":"What one run of this task costs you without the capability, in USD.","minimum":0,"type":"number"},"capability ⟨16 unchanged words⟩ ,"type":"string"},"cost_usd":{"description":"What the run cost you in USD.","minimum":0,"type":"number"},"evaluation_id ⟨32 unchanged words⟩ ,"type":"string"},"latency_ms":{"description":"Wall-clock time of the run in milliseconds.","minimum":0,"type":"integer"},"model ⟨33 unchanged words⟩ ,"type":"string"},"success":{"description":"Whether the task reached its goal.","type":"boolean"},"tokens_in":{"description":"Input tokens the run used.","minimum":0,"type":"integer"},"tokens_out":{"description":"Output tokens the run used.","minimum":0,"type":"integer"},"tool_calls":{"description":"Tool calls the run made.","minimum":0,"type":"integer"},"used_capability":{"description":"Whether the capability was used in this run; false reports a run done without it.","type":"boolean"}},"required":[" ⟨4 unchanged words⟩
⟨22 unchanged words⟩ number"},"constraints":{"additionalProperties":false,"description":"Limits you set for the answer; a capability outside them is reported as blocked, with the reason code.","properties":{"max_cost_usd":{"description":"The highest price in USD you accept; a capability priced above it is blocked (over_max_cost).","minimum":0,"type":"number"},"minimum_confidence":{"description":"The lowest confidence in the measured saving you accept, from 0.5 to 0.999 (default 0.9); below it the answer is blocked (below_minimum_confidence).","maximum":0.999,"minimum":0.5,"type": ⟨84 unchanged words⟩
{"properties":{"limit":{"description":"How many results to return, from 1 to 50 (default 10).","maximum":50,"minimum":1,"type":" ⟨16 unchanged words⟩
Added "resolve_task"; changed the definition of "compare_capabilities", "estimate_roi", "evaluate_capability" and 7 more (10 tools before, 11 now)
Assess2-5two to five named capabilities side by side for the same inputs, with the same method and ak.offer.v1 offer shape as evaluate_capability. Use when: you already hold a shortlist of capability ids and want each one's decision, reasons and offer in one answer. Not for: finding candidates (search_capabilities) or a single capability (evaluate_capability). Input: the capability ids, plus the optional task, model, expected_runs, baseline_cost_estimate_usd, context and constraints that evaluate_capability reads. Output: ak.comparison.v1 with the results ranked by decision, lower-bound net saving, relevance and capability id, each with its reasons and an ak.offer.v1 offer.
Certificate recorded, valid to 2026-12-24
Authorization not required
Unknown → Live
First tool surface recorded: 10 tools (server version 0.1.0)
https://agentickeychain.com/mcp (mcp_streamable_http) — from mcp_registry, with the record
Showing the latest 7 events. The API returns up to 500 and filters by kind: ?kind=tool_surface_changed
Decide whether a registry capability is worth using or buying for your task, from registry-run benchmarks only.ReturnsUse when: you have a task and want a recommend, do_not_recommend or undetermined answer with the arithmetic; name a capability to weigh only that one, or leave it out and the registry weighs its closest matches. Not for: one short answer without the arithmetic (resolve_task), a plain list of matches (search_capabilities), two to five named candidates side by side (compare_capabilities), or the break-even figures of one capability without a task (estimate_roi). Input: a plain task description; give model, expected_runs and baseline_cost_estimate_usd for a money estimate, and context for the conditions the capability declares. Output: ak.evaluation.v1 with decision, reasons, expected effect, price, confidence,reasonsoperator_message,operator_messageevaluation_id, next steps and an ak.offer.v1 offer (status, next_action, economics, the four evidence kinds listed apart).Give model, expected_runs and baseline_cost_estimate_usddo_not_recommendforis amoneynormalestimate.answer.
⟨5 unchanged words⟩ claims (labelled PUBLISHER CLAIMED) and the signedattestations.attestations of one capability. Use when: you want the measurements behind an answer (paired runs, model, fixtures, method) or want to check them yourself. Not for: an answer for your task (evaluate_capability) or artifact hashes and the log proof (get_provenance). Input: a capability id, ak:// URI or slug. Output: ak.benchmark.v1 with the evidence, the publisher claims, the attestations, the method and the limitations; publisher claims never decide anything.
⟨12 unchanged words⟩ trust checks, links and a task-independent ak.offer.v1 offer. Use when: you can name a capability and want its facts, or want to check the conditions an evaluation asked about. Not for: an answer for your task (evaluate_capability) or the full benchmark records (get_benchmark). Input: a capability id, ak:// URI or slug, from a search, an evaluation or a page. Output: ak.capability.v1 with the card fields, parameters, versions, terms, evidence, trust and links.
⟨5 unchanged words⟩ publisher and registry attestations and a transparency-log inclusionproof.proof for one capability. Use when: you want to check that a package is the one the registry published and signed. Not for: benchmark results (get_benchmark) or the card (get_capability). Input: a capability id, ak:// URI or slug. Output: ak.provenance.v1 with the hashes, the attestations, an RFC 9162 inclusion proof against the signed tree head, the trust checks and the keys URL.
⟨14 unchanged words⟩ purchase requires operator approval in the hosted checkout. Use when: your operator wants the price and the checkout link of a capability that is for sale, for example after an evaluation answered recommend. Not for: deciding whether a capability is worth it (evaluate_capability); a free capability needs no quote. Input: the capability id and, if you evaluated first, the evaluation_id, which links the quote to that evaluation. Output: a quote_id and signed claims bound to the capability version, price, currency, terms and expiry, the checkout link and the unlock request; the price comes from the registry record only.
⟨22 unchanged words⟩ fields are accepted; no task text is kept. Use when: you ran a capability, or decided not to, and can say whether the task succeeded. Not for: questions about a capability (get_capability, get_benchmark). Input: the capability id and success; add evaluation_id or purchase_ref to link the report, the optional token, tool-call, latency and cost counts, and failure_reason only with success false. Output: whether the report was stored, its label and, for a failed paid capability, the remedy policy.
Lexical search overSearch capability names, summaries, tasks and use-whenconditions.conditionsRankingbyiskeywords, ranked by lexical relevance only, never by payment. Use when: you want to see which capabilities exist for a topic, or you know part of a name. Not for: deciding whether one fits your task (resolve_task, or evaluate_capability for the full arithmetic) or reading a capability you can already name (get_capability). Input: search words, at most 500 characters; no earlier call is needed. Output: ak.search.v1 with id, name, summary, relevance, state, price and the strongest evidence label of each result, and no decision.
⟨39 unchanged words⟩ the same key is safe and never charges. Use when: your operator completed the hosted checkout and gave you the license key. Not for: a price or a checkout link (get_quote); a free capability downloads from its page. Input: the capability id and the license key, plus quote_id and evaluation_id when you have them. Output: a signed entitlement, a purchase_ref for report_outcome and the download link.
Is there a capability worth using for this task? — Describe a task in plain words and get one compact answer: whether a registry capability is worth using for it, from registry-run benchmarks only, and the next step. Use when: you have a task and no capability name, and want one decision word (use, do_not_use, use_free_alternative, insufficient_evidence, blocked_by_constraints or no_match) with its reason codes. Not for: the full arithmetic, the offer and the operator message (evaluate_capability with the same inputs), a plain list of matches (search_capabilities) or a capability you can already name (get_capability). Input: task, plus option…