Servers, never people
This server keeps statistics about the fediverse for teaching and curiosity. They name servers,
never accounts: what kind of software a server runs, what it exchanges with us, and how reliably. Distinct accounts are
only counted, through a key that is destroyed at the end of each day.
A server we exchange activities with is described once a week, from its public NodeInfo and, when it has one, its
Mastodon instance API. Its location comes from the address we reached, looked up in an offline database: only the
country is shown publicly for small servers, and only the CDN for servers behind one.
The crawler
@if (Model.CrawlerEnabled)
{
The crawler is on on this server.
}
else
{
The crawler is off on this server: it only learns about servers it already exchanges with.
}
When on, it identifies itself as
@Model.UserAgent
It visits one server a minute, each at most once a week, and reads only:
/robots.txt
/.well-known/nodeinfo and the NodeInfo document it points to
/api/v2/instance or /api/v1/instance
/api/v1/instance/peers, to find other servers
It never reads accounts, posts, timelines or directories.
Keeping it out
Add this to your server's robots.txt:
User-agent: @StargazerToken
Disallow: /
If your robots.txt cannot be read because of a server error or a timeout, the crawler stays out too.