Check public crawler policy
Parse robots.txt for major AI search, training, classic search, common crawlers, and Content-Signal declarations at the exact URL path.
AI crawler access check
See which AI crawlers your robots.txt allows, check whether a public page responds, and find practical fixes. Compare GPTBot, ClaudeBot, and other search and training bots in one report.
What it checks
A page’s robots rules, HTTP response, and indexing directives answer different questions. Review them together before changing your crawler policy.
Parse robots.txt for major AI search, training, classic search, common crawlers, and Content-Signal declarations at the exact URL path.
Review HTTP status, redirects, meta robots, X-Robots-Tag, canonical, readable text, and JSON-LD.
Review sitemap discovery, optional llms.txt files, and suggested fixes for the issues found.
Your crawl policy
Set search and training permissions to suit your site. Use your access logs to verify visits; check AI answers separately for citations.
Find accidental blocks that may prevent AI search or retrieval systems from reading public pages.
See which training or retrieval bots you are allowing, then choose a policy intentionally.
Retest the same URL after updating a rule, then compare the result with requests in your server or CDN logs.
Related tools
Start with the full AI crawler access scan, then use the focused pages when the job is llms.txt validation, technical AEO readiness, or AI search visibility prerequisites.
Tool choice
AI crawler access is one layer. Compare it with visibility tracking and llms.txt validation so teams do not treat one passing score or one Content-Signal line as a full AI search strategy.
Best for confirming whether a specific user agent is allowed or blocked at one URL path.
Best after technical blockers are fixed, when teams need prompt sampling, citations, and share-of-answer tracking.
Best for checking AI-readable discovery files, but they still need robots.txt, sitemap, metadata, and readable pages around them.
Best for separating actual crawl access, such as Applebot, from AI-use control tokens, such as Applebot-Extended.
Best for spotting search, AI input, AI training, and optional immediate/reference/full use preferences published alongside crawler rules.