robots.txt Generator
Free, no signup. Set a simple default policy, choose the bot groups to block, then test representative URLs against the generated file before publishing it.
Build your file
Select crawler groups to block. Search crawlers remain allowed unless you explicitly select them.
This non-standard directive is emitted only for Bingbot. Google does not process it; Bing documents 1–20 second values. It is omitted if you block Bingbot.
Runs entirely in your browser — nothing you paste is uploaded or stored. Everything here, including matching and copy output, runs in your browser. Nothing is uploaded. Anonymous run-level outcome counters may be used for aggregate research; URLs, domains, IPs, and identifiers are never included, and no statistic is released below 100 runs.
Generated robots.txt
Blocking ≠ deindexing. A Disallow asks compliant crawlers not to fetch a URL; it does not remove a URL already known to a search engine.
Live match preview
Need a full bot × URL matrix? Open robots.txt Tester →How to use it
- Choose which bot categories should not crawl the site.
- Add your absolute sitemap URL if you have one.
- Paste a few important URLs and a bot name to confirm the intended result.
- Copy the file to the site root as
/robots.txt.
How it works
The output uses separate user-agent groups and Disallow: / rules, so the intention is easy to audit. It uses the same Google-style matcher as the robots.txt Tester, including longest-match-wins behavior.
What this generator does not do
robots.txt is not security, and it is not a deindexing tool. Do not list sensitive paths expecting them to stay secret; the file is public. Use access controls for private material and noindex for index removal.
Frequently asked questions
Does blocking a bot deindex my pages?
No. robots.txt controls crawling, not whether a URL remains indexed. Use a noindex directive on a crawlable page when you need a deindexing signal.
Will every crawler obey this file?
No. robots.txt is a public, voluntary crawl preference, not access control. Use authentication, a WAF, or rate limiting to protect private content.
Why do I need a Sitemap line?
It gives crawlers an explicit absolute URL for your XML sitemap. It is optional, but usually useful for discoverability.
Site passport Local context for this saved site
Local data
Saved targets, named lists, and recent check summaries remain only in this browser.
Rate this tool
Feature requests for Robots Txt Generator
Upvote what you want most. New ideas can be submitted from the floating Feedback menu; requests appear here once approved, and the most-wanted rise to the top.
You won't be emailed about that request anymore.
Loading…
➕ Request a feature
New requests are reviewed before they appear here.
Sobre a ferramenta
Grátis, sem cadastro. Set um simple default policy, escolha o bot groups para bloquear, então test representative URLs against o gerado file antes publicação isso.
Build um claro robots.txt com rastreador presets, ativo Google-style URL correspondente, um sitemap line, e copy-ready saída.
Recursos
- Selecione rastreador groups para bloquear. Pesquisa rastreadores permanecem allowed unless você explicitly selecione them.
- Escolha que bot categories deve não rastreamento o site. Adicione seu absolute sitemap URL if você têm um. Paste um few importante URLs e um bot nome para confirm o pretendido resultado. Copy o file para o site root como /robots.txt .
- Isso gives rastreadores um explicit absolute URL para seu XML sitemap. Isso é optional, but usually useful para discoverability.
- O saída usa separate user-agent groups e Disallow: / rules, so o intention é easy para auditar. Isso usa o mesmo Google-style matcher como o robots.txt Tester , including longest-match-wins comportamento.
Como funciona
O saída usa separate user-agent groups e Disallow: / rules, so o intention é easy para auditar. Isso usa o mesmo Google-style matcher como o robots.txt Tester , including longest-match-wins comportamento.
Limitações
- robots.txt é não security, e isso é não um deindexing ferramenta. Faça não listar sensitive caminhos expecting them para stay secret; o file é público. Usar access controls para privado material e noindex para índice removal.
- Não. robots.txt é um público, voluntary rastreamento preference, não access control. Usar autenticação, um WAF, ou taxa limiting para proteger privado conteúdo.
Perguntas frequentes
faz blockndo a bot deindex my páginas?
não. robots.txt controls rastreamento, não whether a URL remains indexdo. usar a noindex directive em a crawlable página quando você precisa a deindexndo signal.
vai cada crawler obey esta arquivo?
não. robots.txt é a público, voluntary rastrear preference, não acesso controle. usar authenticação, a WAF, ou taxa limitndo para protect private conteúdo.
Por que I precisa a sitemap linha?
ele gives rastreadores um explicit absolute URL para seu XML sitemap. ele é opcional, but normalmente útil para discoverabilidade.