Public site
Leave both path boxes empty to allow everything and just add your sitemap.
Create crawler rules, add your sitemap and download a robots.txt file ready for your website root.
2 lines
Public websites usually allow crawling by default and exclude admin areas, duplicate pages or internal search results.
Leave both path boxes empty to allow everything and just add your sitemap.
Disallow admin, cart or internal search paths to protect crawl budget.
Disallow / for staging sites you do not want compliant crawlers to visit.
Robots.txt controls crawling, not guaranteed indexing. One broad disallow rule can hide important pages from search crawlers.
It asks a crawler not to request URLs that begin with the specified path. An empty Disallow value blocks nothing.
It creates a more specific exception inside a broader blocked path when the crawler supports the directive.
That is a business choice. Blocking training crawlers may limit reuse for model training, but different bots have different purposes and compliance policies.
No. A blocked URL can still be indexed if other pages link to it. Use a noindex directive on a crawlable page when removal from search is the goal.
Rankauto crawls your site, groups issues by severity and shows the pages that need attention.
Get free trial