The agent door · computed, not judged
Machine discoverability
Whether a retrieval system can find and read you at all.
How it is scored
0 to 20
no sitemap, primary content only renders with JavaScript, or crawlers or declared AI agents are blocked outright.
40 to 60
sitemap present and some schema.org markup, but key pages need JavaScript.
80 to 100
crawlers, declared AI agents and real browsers are all admitted, rich schema.org markup, and the main content is readable in raw HTML. llms.txt earns a small credit as declared intent and is never decisive.
What the scanner checks
- Whether a sitemap exists, and whether robots.txt declares it
- How many schema.org types appear on the homepage
- Whether robots.txt blocks assistant crawlers by name
- Whether the page is actually served to requests carrying the named AI crawlers' names, compared against what robots.txt declares
- Whether /llms.txt exists, credited as declared intent and weighted below every access signal, because no major crawler reads it