Robots.txt & Noindex Checker
Find out exactly what search engines and AI crawlers are allowed to see. Fetch and validate any robots.txt, test a specific path as Googlebot, Bingbot, GPTBot, ClaudeBot or PerplexityBot, then scan single pages β or the first pages of a whole sitemap β for hidden noindex tags and canonical mistakes.
How does the Robots.txt Tester work?
The tool fetches your live robots.txt, parses it with the same longest-match rules crawlers use (including * wildcards and $ end-anchors), and shows every user-agent group with its rules. A bot chip row instantly shows whether Googlebot, Bingbot, GPTBot, ClaudeBot, PerplexityBot, Google-Extended and CCBot can access your homepage β critical now that AI search visibility depends on these crawlers. Type any path and pick a bot to get an ALLOWED/BLOCKED verdict with the exact matching rule.
- Fetch & validate: syntax warnings, full-site block detection, missing Sitemap: line alerts.
- Test a path: live per-bot testing with the winning rule shown β no guessing.
- Noindex tab: checks meta robots/googlebot tags, follow status and canonical on any page, or scans the first 12 sitemap URLs at once.
Robots.txt vs noindex β which one hides a page?
They are different tools that people constantly mix up. Robots.txt blocks crawling β but a blocked URL can still appear in Google (without a description) if other sites link to it. Meta noindex removes a page from results β but the crawler must be able to fetch the page to see the tag, so never combine noindex with a robots.txt block. This checker covers both sides so you can see the complete indexation picture, including the accidental sitewide noindex that staging sites love to ship to production.