Robots.txt Checker and Tester
Fetch any site's robots.txt, test whether a URL is allowed for Googlebot, GPTBot, ClaudeBot, or any user agent, and see its sitemaps and AI crawler rules.
How do I test a robots.txt file?
- Enter a domain, such as example.com, and click Fetch robots.txt.
- Choose a user agent from the list or type your own, such as Googlebot or GPTBot.
- Type a path or a full URL to test. The verdict updates as you type and names the rule that decided it.
- Read the AI crawler table, the listed sitemaps, and any ignored lines to see the rest of the file at a glance.
How does robots.txt decide which URLs a crawler can fetch?
The file is split into groups, each starting with one or more User-agent lines followed by Allow and Disallow rules. A crawler picks the group that names it, or the * group if none does, then compares each rule path against the URL path. The longest matching rule wins, and an empty Disallow allows everything. See the robots.txt glossary entry for the full syntax.
Should I block AI crawlers in robots.txt?
It depends on what you want. Blocking a training crawler such as GPTBot or ClaudeBot keeps your content out of future model training, while blocking a search crawler such as OAI-SearchBot or Claude-SearchBot can remove you from AI search answers. The table above shows which of them your file already blocks. Rules for a crawler that is not listed fall through to the * group.
How do I check robots.txt before scraping at scale?
Fetch the file once per domain, cache it, and test each URL against it before you request the page. The web scraping API returns the pages you choose to fetch, and the free sitemap checker shows which pages a site wants crawled.
What can you use a robots.txt tester for?
- Confirming that a new Disallow rule blocks what you expect and nothing else
- Finding out why Googlebot or another crawler cannot reach a page
- Auditing which AI crawlers a site allows before you publish or scrape
- Checking that a robots.txt lists your sitemap
- Reading a competitor's crawl rules before you build a crawler
Frequently asked questions
Is the robots.txt tester free?
What happens when a site has no robots.txt?
Which AI crawlers does the tool check?
Which rules does the tester follow?
Does robots.txt keep a page out of Google or AI answers?
Should I respect robots.txt when I scrape a site?
Is the robots.txt tester free?
Which rules does the tester follow?
What happens when a site has no robots.txt?
Does robots.txt keep a page out of Google or AI answers?
Which AI crawlers does the tool check?
Should I respect robots.txt when I scrape a site?
More free tools
Keep exploring
Twitter Card Validator
Preview how your page looks on X and validate Twitter Card meta tags.
Scrape API (HTML)Sitemap Generator
Valid XML sitemap from any domain, ready for Google Search Console.
Map URLs API- New
Sitemap Checker and Validator
Validate a sitemap and see URL count, lastmod coverage, and errors.
Map URLs API
Ship an agent that actually knows things.
Free tier, 10-minute integration, and the same API powering agents at Mintlify, daily.dev, and Propane. No credit card to start.