robots.txt Generator
Free, no signup. Set a simple default policy, choose the bot groups to block, then test representative URLs against the generated file before publishing it.
Build your file
Select crawler groups to block. Search crawlers remain allowed unless you explicitly select them.
This non-standard directive is emitted only for Bingbot. Google does not process it; Bing documents 1–20 second values. It is omitted if you block Bingbot.
Runs entirely in your browser — nothing you paste is uploaded or stored. Everything here, including matching and copy output, runs in your browser. Nothing is uploaded. Anonymous run-level outcome counters may be used for aggregate research; URLs, domains, IPs, and identifiers are never included, and no statistic is released below 100 runs.
Generated robots.txt
Blocking ≠ deindexing. A Disallow asks compliant crawlers not to fetch a URL; it does not remove a URL already known to a search engine.
Live match preview
Need a full bot × URL matrix? Open robots.txt Tester →How to use it
- Choose which bot categories should not crawl the site.
- Add your absolute sitemap URL if you have one.
- Paste a few important URLs and a bot name to confirm the intended result.
- Copy the file to the site root as
/robots.txt.
How it works
The output uses separate user-agent groups and Disallow: / rules, so the intention is easy to audit. It uses the same Google-style matcher as the robots.txt Tester, including longest-match-wins behavior.
What this generator does not do
robots.txt is not security, and it is not a deindexing tool. Do not list sensitive paths expecting them to stay secret; the file is public. Use access controls for private material and noindex for index removal.
Frequently asked questions
Does blocking a bot deindex my pages?
No. robots.txt controls crawling, not whether a URL remains indexed. Use a noindex directive on a crawlable page when you need a deindexing signal.
Will every crawler obey this file?
No. robots.txt is a public, voluntary crawl preference, not access control. Use authentication, a WAF, or rate limiting to protect private content.
Why do I need a Sitemap line?
It gives crawlers an explicit absolute URL for your XML sitemap. It is optional, but usually useful for discoverability.
Site passport Local context for this saved site
Local data
Saved targets, named lists, and recent check summaries remain only in this browser.
Rate this tool
Feature requests for Robots Txt Generator
Upvote what you want most. New ideas can be submitted from the floating Feedback menu; requests appear here once approved, and the most-wanted rise to the top.
You won't be emailed about that request anymore.
Loading…
➕ Request a feature
New requests are reviewed before they appear here.
Über das Tool
Erstelle eine kopierfertige robots.txt mit getrennten Presets für Such-Crawler, KI-Training, Assistenten und Archiv-Crawler. Ergänze optional einen dokumentierten Bingbot-Crawl-Delay und teste repräsentative URLs gegen die erzeugte Datei mit demselben Google-ähnlichen Matcher wie der robots.txt-Tester.
robots.txt ist eine öffentliche, freiwillige Crawl-Präferenz und weder Sicherheitskontrolle noch Deindexierungswerkzeug.
Funktionen
- Gruppierte User-Agent-Presets mit Allowed-/Blocked-Vorschau und Disallow: / Regeln
- Optionale Sitemap-Zeile sowie dokumentierter Bingbot-Crawl-Delay
- Google-ähnliches Longest-match-Verhalten für Bot- und URL-Testfälle
- Kopieren, Herunterladen und lokale Beispieltests vor dem Veröffentlichen
Funktionsweise
Wähle die zu blockierenden Crawler-Gruppen; Such-Crawler bleiben erlaubt, solange du sie nicht auswählst. Ergänze Sitemap-URL und optionalen Delay, generiere die Datei und teste Beispielpfade. Prüfe die gewinnende Regel und exportiere erst danach die Vorlage für deine Website.
Einschränkungen
- robots.txt schützt keine privaten Inhalte; nutze Authentifizierung, WAF oder Rate-Limits für Zugriffskontrolle.
- Ein Disallow verhindert Crawling, aber nicht zwingend die Indexierung; für Deindexierung brauchst du eine crawlbare noindex-Anweisung.
- Crawler können robots.txt ignorieren und User-Agents fälschen; teste reale Zugriffe zusätzlich in deiner Infrastruktur.
Häufig gestellte Fragen
Ist robots.txt ein Sicherheitsmechanismus?
Nein. Die Datei ist öffentlich und freiwillig. Verwende Authentifizierung, WAF oder Rate-Limits, um private Inhalte zu schützen; trage keine geheimen Pfade in Erwartung von Vertraulichkeit ein.
Entfernt Disallow eine URL aus dem Index?
Nicht zuverlässig. Disallow verhindert das Abrufen, während noindex einer crawlbaren Seite ein Deindexierungssignal gibt. Nutze für Indexentfernung eine erreichbare Seite mit passendem noindex.
Warum brauche ich eine Sitemap-Zeile?
Sie nennt Such-Crawlern die bevorzugte XML-Sitemap und erleichtert Discovery. Sie erlaubt oder blockiert keine URLs und ersetzt weder Canonical- noch Indexierungsprüfung.
Kann ein Crawler die Datei ignorieren?
Ja. robots.txt ist eine freiwillige Präferenz. Für echte Durchsetzung brauchst du Regeln an CDN, WAF oder Server und solltest User-Agent- sowie IP-Nachweise prüfen.