Choose robots.txt or sitemap.xml
Two tabs, two small files most sites need.
Build a robots.txt with the rules your site needs, and a sitemap.xml from a list of pages. Both files are generated as you type, ready to download and upload to your site.
100% private — the files are generated in your browser and nothing is sent to any server.
Two tabs, two small files most sites need.
Presets cover the common cases.
Both files go in the root of your site.
Search engines and other automated crawlers read two files before anything else: robots.txt, which tells them what they may and
may not fetch, and sitemap.xml, which lists the pages worth visiting. Neither is required, but a small site without them leaves crawlers
to guess, and a badly written robots.txt can accidentally hide the whole site from search results. Both files are plain text and easy to get
wrong by hand; this tool builds them from a form.
This file sits at the root of your domain, such as example.com/robots.txt, and lists rules by user agent: which crawler the rule applies to,
which paths are disallowed and which are explicitly allowed. It is a request, not a lock: well-behaved crawlers follow it, but it does not
stop a page from being accessed directly or protect private data. Use it to keep crawlers away from admin areas, search result pages, shopping
carts and duplicate content, and to point to your sitemap. Some site owners now also want to block the crawlers that collect training
data for AI models, such as GPTBot or Google-Extended, which this tool can add with one click, separately from the rules for search engines.
A sitemap lists your pages in a simple XML format so that a crawler can find them without following every link, which helps with new or poorly linked pages. Paste your page paths, or full addresses, and the site address once, and the tool builds the file, adding an optional last modified date, a change frequency and a priority. These last three are hints, not guarantees; search engines increasingly ignore change frequency and priority and rely on their own crawling patterns, but they cost nothing to include and some tools still read them.
A single sitemap file should not list more than 50,000 addresses; larger sites split their pages into several sitemap files listed in a sitemap index, which is beyond what this simple generator produces. Robots.txt rules are advisory for compliant crawlers only. Everything here is built in your browser; nothing is uploaded until you place the files on your own server.
More utilities that also run without leaving your browser.