Robots.txt Generator
A robots.txt file tells search engines and AI crawlers which parts of your website they may visit. This free generator builds that file for you: list the paths you want to keep out of search results, add rules for specific bots, slow polite crawlers down with a crawl-delay, block AI training crawlers with one click, and point every visitor to your sitemap. Everything runs in your browser, and you can copy or download the finished file and upload it to your domain root.
Use Robots.txt Generator
Extra per-bot rule
User-agent: * Allow: /
Read Content
What Is the Robots.txt Generator?
The Robots.txt Generator is a free online tool that writes a valid robots.txt file for your website. A robots.txt file lives at the root of your domain, for example at example.com/robots.txt, and every serious crawler reads it before visiting your pages. It contains simple groups of rules: which bot the rule is for, which paths that bot may or must not request, and optionally where your sitemap lives. Search engines treat these rules as instructions for polite crawling, and most AI crawlers published by large labs honour them too.
You do not need to memorise the syntax. Type the folders you want to keep out, such as /admin/ or /private/, add any special rules for single bots, set a crawl-delay if a crawler hits your small server too hard, tick the AI crawler block if you do not want your content used for model training, and paste your sitemap URL. The tool assembles the groups in the correct order and shows a live preview you can copy or download as a plain robots.txt file.
What This Tool Can Do
- Build the default group for all crawlers with allow-all or a custom disallow list, one path per line.
- Add unlimited per-bot rules, for example a stricter disallow for one bot while everyone else stays allowed.
- Add an optional crawl-delay in seconds for crawlers that request too aggressively.
- Block well-known AI crawlers with one checkbox, including GPTBot, ChatGPT-User, ClaudeBot, anthropic-ai, CCBot, Google-Extended, Bytespider and PerplexityBot.
- Declare your XML sitemap so crawlers discover new pages faster.
- Show a live preview, a one-click copy button, and a download button that saves a ready-to-upload robots.txt file.
- Normalise paths automatically so a missing leading slash never breaks the file.
- Run fully in the browser with no signup, so draft folder names never leave your computer.
Why a Good Robots.txt File Is Useful
Every crawl budget is limited. When a search engine wastes requests on login pages, shopping carts, internal search results, staging copies and temporary folders, it visits your important pages less often. A clean robots.txt file steers crawlers toward the pages that can actually rank and away from the pages that never should. That means faster discovery of new content, fewer junk URLs in the index, and less load on small servers.
It also protects you from embarrassment. Admin panels, client preview folders, debug endpoints and PDF drafts have all been indexed by search engines because no robots.txt rule stopped the crawler. A disallow rule is not access control, but it stops the casual indexing that puts private-looking URLs into public search results. For AI crawlers the value is choice: many site owners are happy to be indexed by search engines but do not want their text reused for model training, and the one-click AI block expresses exactly that preference in the standard format those bots document.
Who This Tool Is For
- Bloggers and small business owners who want a correct robots.txt without learning the syntax.
- Developers setting up a new project who need a sane starting file in under a minute.
- SEO freelancers who write robots.txt files for many clients and want consistent formatting.
- Store owners who must keep carts, checkouts, account pages and filtered faceted URLs out of the index.
- Publishers deciding whether AI crawlers may reuse their articles, and wanting the opt-out in the documented format.
- Anyone migrating a site who needs to add or confirm the sitemap declaration during launch week.
How to Use This Tool
- List the paths to disallow, one per line, for example /admin/, /private/ and /tmp/. Leave the box empty if every crawler may visit every public page.
- Paste your sitemap URL if you have one, for example https://example.com/sitemap.xml. Crawlers use it to find new and updated pages.
- Set a crawl-delay only if a polite crawler overloads a small server. Most large sites leave this empty because major search engines interpret it differently.
- Decide about AI crawlers. Keep the checkbox on to add a disallow-all group for GPTBot, ClaudeBot and the other listed AI bots, or switch it off if you welcome AI traffic and citations.
- Optionally add an extra per-bot rule: type the bot name, choose Allow or Disallow, type the path, and press Add rule. Repeat for as many bots as you need, and remove any row with the Remove button.
- Press Generate robots.txt and read the preview. Check that each bot group shows the paths you intended.
- Press Copy to paste the file into your hosting file manager, or Download robots.txt to save it and upload it to the root of your domain so it loads at yourdomain.com/robots.txt.
- After uploading, open the URL in your browser and test one blocked and one allowed path to confirm the file is live.
Example Output
A typical small site with an admin area, a sitemap, and the AI block switched on produces a file like this:
User-agent: *
Disallow: /admin/
Disallow: /private/
User-agent: GPTBot
Disallow: /
Sitemap: https://example.com/sitemap.xml
The first group speaks to every crawler, the second group speaks only to the named AI bot, and the last line points everyone at the sitemap. Groups are separated by blank lines, exactly as the specification expects.
Limitations You Should Know
Robots.txt is an instruction file, not a security system. Polite crawlers obey it, but malicious scrapers ignore it completely, so never rely on it to hide truly sensitive data. Use login protection and server-side access control for anything that must stay private. The tool also cannot verify your paths: if you disallow /blog/ by accident you will hide your whole blog from search engines, so always double-check every line before uploading.
Crawl-delay is honoured inconsistently across the industry, with some major search engines ignoring it entirely, so treat it as a hint rather than a guarantee. AI crawler blocking works only for bots that respect the standard, and new AI bots appear regularly, so revisit the list a few times per year. Finally, the generator writes the file but cannot upload it for you: the file must be served as plain text at the domain root, and any typo in the filename or location means crawlers will never see it.
What to Do and What Not to Do
Do keep the file short and intentional, with a comment-free list of rules you can explain to a colleague. Do declare the sitemap, because it costs one line and speeds up discovery. Do test after every change with an online robots.txt tester or your search console coverage report. Do keep separate files per subdomain when they serve different content, since each host reads only its own file.
Do not disallow your CSS, JavaScript or image folders unless you have a specific reason, because modern search engines need those files to render pages correctly. Do not use robots.txt to remove a page that is already indexed; instead add a noindex tag or remove the page and let the crawler confirm. Do not block your whole site with Disallow: / and forget to remove it after a staging launch, which is one of the most common self-inflicted SEO outages. Do not paste secret URLs into the file thinking it hides them, because the file itself is public and attackers read it first.
Frequently Asked Questions
FAQ
Comments (0)
Describe the issue you encountered so we can investigate and improve the tool.