WebTooli
Advertisement
750 × 90

Robots.txt Generator

Create a correctly formatted robots.txt file for your website, block AI training bots, and control what search engines can crawl. Free and instant.

💡 How to use: build your rules below, then copy the code. On Blogger, go to Settings, then Crawlers and indexing, turn on Custom robots.txt, and paste it there. On WordPress or most other sites, upload it as a file named exactly robots.txt to your site's root folder, so it loads at yoursite.com/robots.txt.
⚙️ Main Rule (User-agent: *, applies to all crawlers)
+ /search + /admin/ + /wp-admin/ + /cgi-bin/ + /tmp/ + /private/ + /*?* (URLs with parameters)
Google ignores this; some other search engines still honor it.
🤖 Block Specific Crawlers (each checked bot gets its own Disallow: / rule)
AI training & answer bots
SEO & scraper bots
🗺️ Sitemap
For Blogger, common sitemap URLs are /sitemap.xml and /sitemap-pages.xml.
✅ Generated robots.txt
Copied!
Loading...
Note: robots.txt only asks crawlers not to visit a path; well-behaved bots follow it, but it does not password-protect anything.

Free Online Robots.txt Generator – Control How Crawlers See Your Site

A robots.txt file is a small plain-text file placed at the root of your website that tells search engine crawlers and other bots which parts of your site they may visit and which parts they should skip. It is the first thing a well-behaved crawler checks before it starts browsing your pages. Writing this file by hand is simple in theory but easy to get subtly wrong, and a single mistake — like an accidental "Disallow: /" — can hide your entire site from search results. Our free robots.txt generator builds the file for you, correctly formatted every time.

Start from a ready-made preset, such as Allow Everything, Block Everything, or the Blogger preset, or build your own rules from scratch. Add the paths you want to keep away from crawlers, such as admin folders, internal search pages, or duplicate content, using the quick-add buttons for common paths. With growing concern over AI companies scraping content to train their models, this tool also includes one-click checkboxes to block AI crawlers like GPTBot, ClaudeBot, Google-Extended, CCBot, and PerplexityBot, along with SEO scraper bots like AhrefsBot and SemrushBot, while still allowing normal search engines like Google and Bing to index your site.

Why Your Site Needs a Robots.txt File

Without a robots.txt file, crawlers will generally try to access every page on your site, which can waste crawl budget on low-value pages like admin panels or search results, leaving your important content crawled less often. A well-configured file directs crawlers toward the pages that matter and away from the ones that don't. It gives you a simple, standard way to signal which AI bots you don't want scraping your content for training data. It also lets you point crawlers directly to your sitemap, ensuring new posts are discovered quickly. And it can prevent accidental indexing of internal pages, staging areas, and files that were never meant to be public, which protects both your SEO and your visitors' privacy.

Understanding User-agent, Disallow, and Allow Rules

Every rule in a robots.txt file is built around three directives. User-agent identifies which crawler the rule applies to — using "*" means the rule applies to all crawlers, while naming a specific bot (like GPTBot) creates an exception just for that crawler. Disallow tells the crawler not to visit a path, and Allow carves out an exception inside a disallowed folder — this is why you often see both together, for example disallowing /wp-admin/ but allowing /wp-admin/admin-ajax.php, since WordPress needs that one file to be reachable. Paths are matched by prefix: "Disallow: /search" blocks /search, /search?q=hello, and /search-results, all at once. Wildcards are supported by Google and Bing — "/*?*" blocks all URLs that contain a query string, which is useful for stopping duplicate-content crawl traps. One important limit: robots.txt only asks well-behaved crawlers to stay away. It does not password-protect anything, and it does not remove already-indexed pages from search results — for that, you need a "noindex" meta tag on the page itself, or password protection. And because the file must sit at your domain root to be found, on Blogger you do not upload it at all: you paste the code into Settings → Crawlers and indexing → Custom robots.txt, and Blogger serves it automatically at yoursite.com/robots.txt.

Free, Instant, and No Signup

This generator runs entirely in your browser, updates the code live as you build your rules, and requires no account or installation. Once you're happy with the result, copy it or download it as a ready-to-use robots.txt file. It works on Windows, Mac, Android, and iPhone, with no limit on how many times you generate or how complex your rules are.

Frequently Asked Questions

Does Disallow in robots.txt remove a page from Google search results?
Not by itself. Disallow only tells crawlers not to visit a page, but if other sites link to that page, Google can still show it in search results without a description, since it was never crawled. To fully keep a page out of search results, use a "noindex" meta tag on the page itself, or password-protect it, rather than relying on robots.txt alone.
Will robots.txt stop AI bots from scraping my content?
It will stop any AI crawler that chooses to respect the standard, and major ones like GPTBot, ClaudeBot, and Google-Extended generally do follow it. However, robots.txt is a voluntary request, not a security lock, so a bot that ignores the rules can still technically access public pages. It remains the standard, widely respected first step for signaling that you don't want your content used for AI training.
Where do I put my robots.txt file, and does it work for Blogger?
A robots.txt file must sit at the root of your domain, for example yoursite.com/robots.txt, not inside a subfolder. On most platforms, this means uploading a file literally named robots.txt to your site's root directory. On Blogger, you don't upload a file at all; instead, go to Settings, then Crawlers and indexing, turn on Custom robots.txt, and paste the generated code directly into that box.
What is the difference between Disallow and Allow?
Disallow tells a crawler not to visit a path. Allow creates an exception inside a disallowed path — for example, if you disallow /wp-admin/ but need the admin-ajax.php file inside it to remain accessible, you add an "Allow: /wp-admin/admin-ajax.php" line. Allow rules always override Disallow rules for the exact path they match.
Kya robots.txt mein Disallow lagane se page Google search se hat jata hai?
Sirf isse nahi. Disallow crawlers ko sirf ye batata hai ke wo page visit na karein, lekin agar koi aur site us page ko link karti hai, to Google bina description ke bhi wo page dikha sakta hai, kyunke wo crawl hi nahi hua tha. Page ko search se poori tarah hatane ke liye page par "noindex" meta tag lagayein, ya password protect karein, sirf robots.txt par bharosa na karein.
Kya robots.txt AI bots ko mera content scrape karne se rok sakta hai?
Ye un AI crawlers ko rok sakta hai jo iss standard ko follow karte hain, aur GPTBot, ClaudeBot aur Google-Extended jaisay bade bots aam tor par isse follow karte hain. Lekin robots.txt ek marzi se maani jane wali guzarish hai, security lock nahi, is liye jo bot ise nazar-andaz kare wo public pages tak phir bhi pohanch sakta hai. Phir bhi, ye AI training ke liye content na chahne ka pehla aur sab se maqbool tareeqa hai.
Robots.txt file kahan lagani chahiye, aur kya ye Blogger par kaam karti hai?
Robots.txt file hamesha domain ke root par honi chahiye, jaise yoursite.com/robots.txt, kisi subfolder mein nahi. Zyadatar platforms par, iska matlab hai ke robots.txt naam ki file root directory mein upload ki jaye. Blogger par file upload nahi karni hoti; iske bajaye Settings mein jayein, phir Crawlers and indexing, Custom robots.txt on karein, aur generate ki gayi code seedha us box mein paste kar dein.
Disallow aur Allow mein kya farq hai?
Disallow crawler ko batata hai ke kisi path par na jaye. Allow ek disallowed path ke andar exception banata hai — misal ke tor par agar aap /wp-admin/ disallow karte hain lekin uske andar admin-ajax.php accessible rakhna chahte hain, to aap "Allow: /wp-admin/admin-ajax.php" line add karte hain. Allow rules hamesha exact match par Disallow ko override karte hain.
← Back to all tools
Advertisement
750 × 90

WebTooli — Free Online Tools & Web Utilities

A curated collection of free web tools to boost your productivity — organized by category, running entirely in your browser.

Image Tools

Advertisement
750 × 90

PDF Tools

Text Tools

Advertisement
750 × 90

SEO Tools

Calculator Tools

Converter Tools

Advertisement
750 × 90

Developer Tools

Other Tools