[SYS.CRAWL // ROBOTS + SITEMAP]TELL CRAWLERS
TELL CRAWLERS
EXACTLY WHAT TO DO.
Both crawl files, generated and downloadable in one pass.
Set your domain, blocked paths and page list. We emit a valid robots.txt — with explicit allow rules for AI answer engines — and a matching sitemap.xml you can drop into your site root.
[ CONFIG ]
[ ROBOTS.TXT ]
# robots.txt — generated with cyberspaghetti.io/robots-sitemap User-agent: * Allow: / Disallow: /admin Disallow: /cart Disallow: /thank-you # AI search engines and answer engines User-agent: GPTBot Allow: / User-agent: OAI-SearchBot Allow: / User-agent: ClaudeBot Allow: / User-agent: PerplexityBot Allow: / User-agent: Google-Extended Allow: / User-agent: Bytespider Allow: / User-agent: CCBot Allow: / Sitemap: https://example.com/sitemap.xml
[ SITEMAP.XML ]
<?xml version="1.0" encoding="UTF-8"?>
<urlset xmlns="http://www.sitemaps.org/schemas/sitemap/0.9">
<url>
<loc>https://example.com/</loc>
<lastmod>2026-08-18</lastmod>
<changefreq>weekly</changefreq>
<priority>1.0</priority>
</url>
<url>
<loc>https://example.com/about</loc>
<lastmod>2026-08-18</lastmod>
<changefreq>weekly</changefreq>
<priority>0.8</priority>
</url>
<url>
<loc>https://example.com/services</loc>
<lastmod>2026-08-18</lastmod>
<changefreq>weekly</changefreq>
<priority>0.8</priority>
</url>
<url>
<loc>https://example.com/pricing</loc>
<lastmod>2026-08-18</lastmod>
<changefreq>weekly</changefreq>
<priority>0.8</priority>
</url>
<url>
<loc>https://example.com/contact</loc>
<lastmod>2026-08-18</lastmod>
<changefreq>weekly</changefreq>
<priority>0.8</priority>
</url>
</urlset>