Articles ยท SEO Tools

Compare robots.txt to a verified AI crawler registry

Publishers want clarity on how AI crawlers read robots.txt, but inventing user-agents is worse than listing none. The registry must cite official documentation, include last-verified dates, and state clearly that compliance is not guaranteed.

Why this is worth doing in the browser

DevNestro’s AI Crawler Checker keeps a central registry (Name, UserAgent, Organization, Purpose, Source, LastVerifiedDate, Enabled) drawn from OpenAI, Anthropic, Google, Apple, and Perplexity docs. Paste robots.txt or fetch /robots.txt safely, then see Allowed/Blocked/No Matching Rule per crawler for a path. Copy explains robots.txt is preference, not hard access control.

How to use the tool

Paste or fetch robots.txt, pick a path, review each crawler’s verdict and source link, then adjust Robots.txt Generator output. Re-verify registry dates when vendors change bots.

  1. Do not invent crawlers without official sources.
  2. Google-Extended and Applebot-Extended are control tokens, not separate log UAs.
  3. User-initiated fetchers may ignore robots.txt.
  4. Preferences ≠ guarantees.

Privacy

Paste mode is local; fetch mode uses the safe public fetcher for /robots.txt only.

Check AI crawler preferences against a verified registry before you publish a new robots.txt.

Open Ai Crawler Checker

An unhandled error has occurred. Reload Dismiss

Rejoining the server...

Rejoin failed... trying again in seconds.

Failed to rejoin.
Please retry or reload the page.

The session has been paused by the server.

Failed to resume the session.
Please retry or reload the page.