Endpoints: 28,729MCP servers: 18,413Payout addresses: 2,070Paid calls: 1,521Letters: 13Defects: 1,321counted 4 min ago
teppi

MCP serverdev.agentreliability/agent-reliability

Testing, benchmarking and auditing autonomous AI agents — methods, harnesses, evidence
UNRATEDActivestreamable-httpagentreliability.dev

Overview

Score?
UNRATED 0.672
of what a free look can see, on 30 looks
Looks
36
last 42 min ago
Tools
9
changed 27 days ago

More info

URL
agentreliability.dev/mcp?via=manifest
streamable-http
Says it is
dev.agentreliability/agent-reliability 0.9.2
protocol 2025-06-18
In the record since
32 days ago

Among servers18,413 with a card

0median 0.606 · this server 0.672 · highest on record 0.8561

Toolsfrom sha256:bad192a4dc…01854f · +0 −0 27 days ago

The tools this server lists, read out of the definition it returned
ToolSchema
answer
Answer a question from the corpus, or refuse. Returns only the claims that bear on the question, each with the sources it cites and its editorial confidence. When the corpus cannot
input · output
compare
Two to six knowledge objects side by side: their cards, every indexed attribute as a matrix (the same fields api/index.json publishes, null where an object does not say), the tags
input · output
get_entity
Fetch one knowledge object by id, with its claims and the sources each claim cites. Use this once search, answer or get_topic has given you an id. An unknown id is not a dead end:
input · output
get_latest
Most recently verified knowledge objects (freshness signal). Use this to judge how current the corpus is, or to see what changed since you last read it. It ranks by verification da
input · output
get_overview
Corpus overview: what this instance knows, counts by type, published tags, freshness. Use this first when you land here and do not yet know whether this corpus can answer your ques
input · output
get_related
Graph neighbours of an object: outgoing and incoming relations, each with its relation type. PAGED: 25 relations by default, 200 at most, and a response budget of about 64 KB per c
input · output
get_sources
The instance's source registry — each entry with its evidence tier, reliability and access date. PAGED: 25 entries by default, 200 at most, and a response budget of about 64 KB per
input · output
get_topic
List the knowledge objects carrying a tag (topics are content-backed tags). PAGED: 25 objects by default, 200 at most, and a response budget of about 64 KB per call — a page over b
input · output
search
Full-text search over the knowledge graph. Matching ignores accents and apostrophes, so query in the user's own words; every hit carries the fields it matched and a score. BM25 rel
input · output
Verify it yourselfnpx teppi-check https://agentreliability.dev/mcp?via=manifestcurl -s https://api.teppi.xyz/v1/trust/mcp/mcs_01M1FZ2A5HYAS4TZRC3RW7A0VY