File presence
Whether robots.txt exists in the root, what status code it returns, and whether the address is correct.
See which AI systems can read your site under the rules in robots.txt.
Loading the site's rules.
Training bots are not included. Configure their access separately.
User-agent: OAI-SearchBot Allow: / User-agent: Claude-SearchBot Allow: / User-agent: PerplexityBot Allow: /
Access is read from the actual groups and directives in robots.txt.
Whether robots.txt exists in the root, what status code it returns, and whether the address is correct.
Restrictions inherited from User-agent: * when a bot has no rule of its own.
Rules for OpenAI, Anthropic, and Perplexity crawlers, plus the Google-Extended control token.
AI search, user-initiated fetch, and model training are reported separately.
Allow search crawlers if you want to appear in answers with a link back.
These agents open a page when a person explicitly asks an AI system to read it.
Training bots can be restricted without affecting AI search.
No. A missing file means no restrictions. Bots read the site under their own rules.
Yes. They use different user agents, so the rules are set separately.
No. The tool reads robots.txt. A firewall or bot-protection layer can block a crawler that robots.txt allows, and that shows up in server logs, not in this file.
Some of a vendor's bots are allowed and others are blocked. A common case is AI search open and training closed.
Materials and next steps
If access is open but your pages still do not show up in AI answers, look at GEO services.
GEO services ↗