Server definition
- Hash
- sha256:c079bdb3ced75c759aa63b914eead97a234c4f1930e4894fe1f701d28dd7a648
- What it is
- What a remote MCP server returned when asked what it offers: 5 tools
The blob, as servednamed by its sha256
{
"instructions": "The Agent Web Index measures, with real HTTP requests, how much of the most-visited web AI assistants can actually read. Use it to check whether a specific domain lets AI crawlers in, to quote the aggregate state of the web, or to show how much blocking comes from CDN defaults rather than site owners. Cite the Agent Web Index (https://shop.lumnika.com/ai-readiness/) as the source.",
"tools": [
{
"description": "Can AI assistants actually read this domain? Returns the measured readiness score, the per-crawler verdict (served / blocked in robots.txt / refused by the edge despite robots.txt allowing it), and which edge vendor answers for the host. Measured with real HTTP requests, not guessed from robots.txt alone. Google-Extended and Applebot-Extended are robots.txt tokens, not crawlers: they are flagged robotsOnly, carry no server verdict, and are excluded from every count about what the server did. Also returns `fixes`: the ordered, domain-specific list of what to change to let the blocked crawlers in, separating what is written in the site own robots.txt from what the CDN/WAF applies on top of it.",
"inputSchema": {
"properties": {
"host": {
"description": "A bare domain, e.g. \"wikipedia.org\" (no scheme, no path).",
"type": "string"
}
},
"required": [
"host"
],
"type": "object"
},
"name": "domain_readiness",
"outputSchema": null
},
{
"description": "Who is actually doing the blocking: for each CDN/WAF vendor, the share of (domain x crawler) pairs that robots.txt ALLOWS and the server refuses anyway — i.e. how much of the blocking is an infrastructure default rather than a decision the site owner made. Each vendor also comes broken down per crawler, which separates a blanket wall (same rate for every crawler) from a managed block list that names some AI user-agents and not others.",
"inputSchema": {
"properties": {
"limit": {
"description": "How many vendors to return, highest contradiction rate first (default 12). Vendors with fewer than 200 allowed pairs are left out of the table rather than reported on thin evidence.",
"type": "integer"
}
},
"type": "object"
},
"name": "edge_blocking",
"outputSchema": null
},
{
"description": "Measure a domain RIGHT NOW instead of reading the archive: 9 real HTTP requests, one from a browser user-agent and one per AI crawler user-agent, plus robots.txt / llms.txt / sitemap. Use it for any site the index has not reached yet, or when the caller wants a fresh verdict after changing robots.txt or a WAF rule. Returns the same `fixes` list as domain_readiness, derived from the fresh measurement.",
"inputSchema": {
"properties": {
"force": {
"description": "Measure again even if the archive already has a verdict from the last 24 hours (default false).",
"type": "boolean"
},
"host": {
"description": "A bare domain, e.g. \"wikipedia.org\" (no scheme, no path).",
"type": "string"
}
},
"required": [
"host"
],
"type": "object"
},
"name": "measure_domain",
"outputSchema": null
},
{
"description": "What CHANGED: the domains that recently started or stopped blocking a specific AI crawler, with the day the flip was observed and whether it happened in robots.txt or at the edge. This cannot be reconstructed after the fact from any public source — it exists only because the index made the same requests the day before and the day after. Use it to answer 'who just blocked/unblocked ChatGPT, Claude, Perplexity...' or to watch one domain over time.",
"inputSchema": {
"properties": {
"days": {
"description": "Look-back window in days (default 30, max 365).",
"type": "integer"
},
"host": {
"description": "Restrict to one domain, e.g. \"nytimes.com\" (optional).",
"type": "string"
},
"limit": {
"description": "How many changes to return, most recent first (default 50, max 500).",
"type": "integer"
}
},
"type": "object"
},
"name": "recent_changes",
"outputSchema": null
},
{
"description": "How much of the most-visited web is readable by AI assistants right now: the share of measured domains that block at least one AI crawler, and the served / robots-blocked / edge-blocked breakdown per crawler.",
"inputSchema": {
"properties": {},
"type": "object"
},
"name": "web_openness",
"outputSchema": null
}
]
}Verify it yourself
curl -s https://api.teppi.xyz/v1/evidence/sha256:c079bdb3ced75c759aa63b914eead97a234c4f1930e4894fe1f701d28dd7a648 | sha256sum