Fetch any site's robots.txt and check whether a path is crawlable by Googlebot — or any user-agent you specify.
Even an allowed path can get rate-limited, fingerprinted or geo-blocked. Rainproxy routes your crawler through real residential and mobile IPs that pass Cloudflare, Akamai and DataDome cleanly.
We fetch robots.txt, parse every group exactly the way Google does, and tell you whether your target path is fair game.
Server-side request grabs the live robots file — no CORS, no caching surprises, no spoofed responses.
User-agent groups, allow/disallow precedence and sitemaps are read using Google's longest-match logic.
Tells you Allow vs Disallow for Googlebot, Bingbot or any custom UA — with the exact matching rule.
Run before launching a scraper or SEO crawler to avoid wasted budget on disallowed paths.
Surfaces every Sitemap: directive so you can point your indexer or scraper at the right URLs.
We always show what the site asks crawlers to do — respecting robots.txt is good citizenship and good SEO.