SEO Tool Tool updated 2 hours ago

Robots.txt Generator

A robots.txt file tells search engines and AI crawlers which parts of your website they may visit. This free generator builds that file for you: list the paths you want to keep out of search results, add rules for specific bots, slow polite crawlers down with a crawl-delay, block AI training crawlers with one click, and point every visitor to your sitemap. Everything runs in your browser, and you can copy or download the finished file and upload it to your domain root.

Always Free Easy to use Instant results Private and secure

Use Robots.txt Generator

Each line becomes a Disallow rule under User-agent: *. Leave empty to allow everything.

Extra per-bot rule


User-agent: *
Allow: /

Read Content

Learn how the tool works and when to use it.

What Is the Robots.txt Generator?

The Robots.txt Generator is a free online tool that writes a valid robots.txt file for your website. A robots.txt file lives at the root of your domain, for example at example.com/robots.txt, and every serious crawler reads it before visiting your pages. It contains simple groups of rules: which bot the rule is for, which paths that bot may or must not request, and optionally where your sitemap lives. Search engines treat these rules as instructions for polite crawling, and most AI crawlers published by large labs honour them too.

You do not need to memorise the syntax. Type the folders you want to keep out, such as /admin/ or /private/, add any special rules for single bots, set a crawl-delay if a crawler hits your small server too hard, tick the AI crawler block if you do not want your content used for model training, and paste your sitemap URL. The tool assembles the groups in the correct order and shows a live preview you can copy or download as a plain robots.txt file.

What This Tool Can Do

  • Build the default group for all crawlers with allow-all or a custom disallow list, one path per line.
  • Add unlimited per-bot rules, for example a stricter disallow for one bot while everyone else stays allowed.
  • Add an optional crawl-delay in seconds for crawlers that request too aggressively.
  • Block well-known AI crawlers with one checkbox, including GPTBot, ChatGPT-User, ClaudeBot, anthropic-ai, CCBot, Google-Extended, Bytespider and PerplexityBot.
  • Declare your XML sitemap so crawlers discover new pages faster.
  • Show a live preview, a one-click copy button, and a download button that saves a ready-to-upload robots.txt file.
  • Normalise paths automatically so a missing leading slash never breaks the file.
  • Run fully in the browser with no signup, so draft folder names never leave your computer.

Why a Good Robots.txt File Is Useful

Every crawl budget is limited. When a search engine wastes requests on login pages, shopping carts, internal search results, staging copies and temporary folders, it visits your important pages less often. A clean robots.txt file steers crawlers toward the pages that can actually rank and away from the pages that never should. That means faster discovery of new content, fewer junk URLs in the index, and less load on small servers.

It also protects you from embarrassment. Admin panels, client preview folders, debug endpoints and PDF drafts have all been indexed by search engines because no robots.txt rule stopped the crawler. A disallow rule is not access control, but it stops the casual indexing that puts private-looking URLs into public search results. For AI crawlers the value is choice: many site owners are happy to be indexed by search engines but do not want their text reused for model training, and the one-click AI block expresses exactly that preference in the standard format those bots document.

Who This Tool Is For

  • Bloggers and small business owners who want a correct robots.txt without learning the syntax.
  • Developers setting up a new project who need a sane starting file in under a minute.
  • SEO freelancers who write robots.txt files for many clients and want consistent formatting.
  • Store owners who must keep carts, checkouts, account pages and filtered faceted URLs out of the index.
  • Publishers deciding whether AI crawlers may reuse their articles, and wanting the opt-out in the documented format.
  • Anyone migrating a site who needs to add or confirm the sitemap declaration during launch week.

How to Use This Tool

  1. List the paths to disallow, one per line, for example /admin/, /private/ and /tmp/. Leave the box empty if every crawler may visit every public page.
  2. Paste your sitemap URL if you have one, for example https://example.com/sitemap.xml. Crawlers use it to find new and updated pages.
  3. Set a crawl-delay only if a polite crawler overloads a small server. Most large sites leave this empty because major search engines interpret it differently.
  4. Decide about AI crawlers. Keep the checkbox on to add a disallow-all group for GPTBot, ClaudeBot and the other listed AI bots, or switch it off if you welcome AI traffic and citations.
  5. Optionally add an extra per-bot rule: type the bot name, choose Allow or Disallow, type the path, and press Add rule. Repeat for as many bots as you need, and remove any row with the Remove button.
  6. Press Generate robots.txt and read the preview. Check that each bot group shows the paths you intended.
  7. Press Copy to paste the file into your hosting file manager, or Download robots.txt to save it and upload it to the root of your domain so it loads at yourdomain.com/robots.txt.
  8. After uploading, open the URL in your browser and test one blocked and one allowed path to confirm the file is live.

Example Output

A typical small site with an admin area, a sitemap, and the AI block switched on produces a file like this:

User-agent: *
Disallow: /admin/
Disallow: /private/

User-agent: GPTBot
Disallow: /

Sitemap: https://example.com/sitemap.xml

The first group speaks to every crawler, the second group speaks only to the named AI bot, and the last line points everyone at the sitemap. Groups are separated by blank lines, exactly as the specification expects.

Limitations You Should Know

Robots.txt is an instruction file, not a security system. Polite crawlers obey it, but malicious scrapers ignore it completely, so never rely on it to hide truly sensitive data. Use login protection and server-side access control for anything that must stay private. The tool also cannot verify your paths: if you disallow /blog/ by accident you will hide your whole blog from search engines, so always double-check every line before uploading.

Crawl-delay is honoured inconsistently across the industry, with some major search engines ignoring it entirely, so treat it as a hint rather than a guarantee. AI crawler blocking works only for bots that respect the standard, and new AI bots appear regularly, so revisit the list a few times per year. Finally, the generator writes the file but cannot upload it for you: the file must be served as plain text at the domain root, and any typo in the filename or location means crawlers will never see it.

What to Do and What Not to Do

Do keep the file short and intentional, with a comment-free list of rules you can explain to a colleague. Do declare the sitemap, because it costs one line and speeds up discovery. Do test after every change with an online robots.txt tester or your search console coverage report. Do keep separate files per subdomain when they serve different content, since each host reads only its own file.

Do not disallow your CSS, JavaScript or image folders unless you have a specific reason, because modern search engines need those files to render pages correctly. Do not use robots.txt to remove a page that is already indexed; instead add a noindex tag or remove the page and let the crawler confirm. Do not block your whole site with Disallow: / and forget to remove it after a staging launch, which is one of the most common self-inflicted SEO outages. Do not paste secret URLs into the file thinking it hides them, because the file itself is public and attackers read it first.

Frequently Asked Questions

FAQ

Short answers for the questions users ask before trusting the result.

Upload it to the root of your domain so it opens at yourdomain.com/robots.txt. A file placed in a subfolder will not be found, because crawlers only check the root of each host.

It prevents crawling, which usually prevents indexing of new pages, but pages that are already indexed may stay visible until the search engine recrawls them. For reliable removal, combine the disallow rule with a noindex tag or delete the page.

No. Search crawlers and AI training crawlers use different user-agent names. Blocking names such as GPTBot or CCBot does not block Googlebot, so normal search indexing continues while AI reuse is refused.

Usually not. Most sites should leave it empty. Set one only when server logs show a polite crawler requesting pages faster than your small server can handle.

The longest matching path wins. A specific Allow for /blog/free/ beats a general Disallow for /blog/, which is why precise per-bot rules are so useful.

Yes. The format is the same everywhere. WordPress users can paste the output into their SEO plugin file editor or upload it by FTP, while static site users place it next to the homepage file before deploying.

No. Generation happens in your browser with JavaScript. Folder names and sitemap URLs never leave your computer until you upload the finished file yourself.

Comments (0)

For user discussion, edge cases, and follow-up tips.
Leave a comment
Your comment will appear publicly after submission.
No comments yet. Be the first to comment!
Let Us Know!
Report an issue with this tool

Describe the issue you encountered so we can investigate and improve the tool.

What helps most
Include the input, expected result, actual result, browser/device, and a screenshot when helpful.
The screenshot button captures the page in the browser and attaches the generated image to this form.
Full-page screenshot
Captured in your browser with newisty.
No screenshot captured yet.
Tool Report