robots.txt Generator
Free, no signup. Set a simple default policy, choose the bot groups to block, then test representative URLs against the generated file before publishing it.
Build your file
Select crawler groups to block. Search crawlers remain allowed unless you explicitly select them.
This non-standard directive is emitted only for Bingbot. Google does not process it; Bing documents 1–20 second values. It is omitted if you block Bingbot.
Runs entirely in your browser — nothing you paste is uploaded or stored. Everything here, including matching and copy output, runs in your browser. Nothing is uploaded. Anonymous run-level outcome counters may be used for aggregate research; URLs, domains, IPs, and identifiers are never included, and no statistic is released below 100 runs.
Generated robots.txt
Blocking ≠ deindexing. A Disallow asks compliant crawlers not to fetch a URL; it does not remove a URL already known to a search engine.
Live match preview
Need a full bot × URL matrix? Open robots.txt Tester →How to use it
- Choose which bot categories should not crawl the site.
- Add your absolute sitemap URL if you have one.
- Paste a few important URLs and a bot name to confirm the intended result.
- Copy the file to the site root as
/robots.txt.
How it works
The output uses separate user-agent groups and Disallow: / rules, so the intention is easy to audit. It uses the same Google-style matcher as the robots.txt Tester, including longest-match-wins behavior.
What this generator does not do
robots.txt is not security, and it is not a deindexing tool. Do not list sensitive paths expecting them to stay secret; the file is public. Use access controls for private material and noindex for index removal.
Frequently asked questions
Does blocking a bot deindex my pages?
No. robots.txtA plain-text file at the root of a host that tells crawlers which URLs they may and may not request. It controls crawling, not indexing — a blocked URL can still be indexed if it's linked from elsewhere. controls crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor., not whether a URL remains indexed. Use a noindexNoindex is a directive that tells search engines to keep a page out of their index, so it won't appear in search results. It works only on pages a crawler can actually fetch — a page blocked in robots.txt can never be noindexed. directive on a crawlable page when you need a deindexing signalDeindexing means getting a URL to stop appearing in Google's search results. There's no single delete button — the right method depends on whether you own the page, whether removal is temporary or permanent, and whether the content should still exist..
Will every crawler obey this file?
No. robots.txt is a public, voluntary crawl preference, not access control. Use authentication, a WAF, or rate limiting to protect private content.
Why do I need a Sitemap line?
It gives crawlers an explicit absolute URL for your XML sitemapAn XML sitemap is a UTF-8 file listing the canonical URLs on your site (with optional lastmod) so search engines can discover and prioritize them. It's a discovery and diagnostic aid, not a guarantee of indexing — and Google ignores its priority and changefreq tags.. It is optional, but usually useful for discoverability.
Site passport Local context for this saved site
Local data
Saved targets, named lists, and recent check summaries remain only in this browser.
Rate this tool
Feature requests for Robots Txt Generator
Upvote what you want most. New ideas can be submitted from the floating Feedback menu; requests appear here once approved, and the most-wanted rise to the top.
You won't be emailed about that request anymore.
Loading…
➕ Request a feature
New requests are reviewed before they appear here.
Common issues & how to fix them
- Warning Generated policy blocks a selected AI crawler Fix: Remove the selected crawler’s Disallow rule, or narrow it to only the paths that should remain unavailable to that AI crawler.
- Warning Generated policy blocks a search crawler Fix: Remove or narrow the Googlebot/Bingbot Disallow rule so pages intended for organic search are crawlable.
- Warning Sensitive path is not disallowed Fix: Add explicit Disallow rules for non-public admin, staging, cart, or internal-search paths, while protecting truly sensitive data with authentication rather than robots.txt.
- Information Generated file has no sitemap declaration Fix: Add an absolute Sitemap: https://…/sitemap.xml declaration pointing to the canonical, publicly fetchable sitemap or sitemap index.
- Warning Generated groups contain conflicting rules Fix: Combine rules for the same user-agent into one group and resolve overlapping Allow/Disallow paths to express one intentional policy.