THE FOMITE

Census / live & auto-updating

The agent-facing web — a census

The rest of this site measures the agents that visit us. This page turns that around: FomiteBot, a small, disclosed, honestly-identified crawler, visits sites that already reached out to agents — their own +https:// calling card, left in their User-Agent, is how we found them — and records, structurally, which agent-facing surfaces they actually expose. See /methodology for exactly how a probe is scored.

15Sites crawled
6.7%Expose any MCP surface
2Block our datacenter IP
13.3%Blocked, of crawled

1. Which surfaces sites actually expose

“Present” means a probe of that exact path returned a structural presence verdict — not a 200. A homepage re-served at every path (a soft-404) is scoredabsent, not present; a robots.txt Disallow is scored disallowed, never routed around.

SurfaceSites exposing it% of crawled
llms.txt426.7%
ai.txt00.0%
.well-known/mcp.json00.0%
.well-known/ai-plugin.json16.7%
.well-known/agent.json00.0%
/mcp16.7%
security.txt213.3%
openapi.json00.0%

2. How robots.txt treats us

fail-closed means the site's own robots.txt could not be fetched or parsed cleanly and we treated that as a refusal rather than assume permission; an allow-all after a 404 means no robots.txt existed at all.

robots.txt outcomeSites
Parsed normally10
No robots.txt (allow all)0
No applicable group0
Fail-closed (fetch/parse failed)5
Other0

3. Site by site

Every crawled site, most recently crawled first, and what it exposes.

HostSourceBlocks DC?llms.txtai.txt.well-known/mcp.json.well-known/ai-plugin.json.well-known/agent.json/mcpsecurity.txtopenapi.json
netapi.comseedyes
aive.globalseed
mcpbeat.comseedblocks
support.anthropic.comseed
www.baidu.comseed
openai.comseedblocks
api.slack.comseedyesyesyes
awario.comseed
code.google.comseed
roninforge.orgseedyes
sec.sqrx.ioseed
www.domainarrivals.comseed
github.comseedyesyesyes
immersivecommons.comseed
smithery.aiseed

Cite this

Every figure here is recomputed live and is CC0. Take the JSON at /api/agent-web, read how it is measured (with the weaknesses stated) at /methodology, and read about the crawler itself, how to verify it, and how to block it at /fomitebot. If your site is listed here and something is wrong, tell us and we will fix it and say that we did.


The Fomite is a field station on the agentic web. Start here · FomiteBot, the crawler · the Observatory.