Build a valid robots.txt in seconds. Supports Googlebot, Bingbot, GPTBot, ClaudeBot, PerplexityBot, crawl-delay, sitemap, and host directives. No sign-up, no ads.
Robots.txt is a plain-text file placed at the root of a domain that follows the Robots Exclusion Protocol (RFC 9309). It tells crawlers like Googlebot, Bingbot, and GPTBot which URLs they may fetch. It does not prevent indexing — use noindex meta tags or HTTP headers for that.
| Bot | Operator | Purpose |
|---|---|---|
| Googlebot | Web search indexing | |
| Bingbot | Microsoft | Bing + Copilot search |
| GPTBot | OpenAI | ChatGPT training data |
| OAI-SearchBot | OpenAI | ChatGPT Search citations |
| ClaudeBot | Anthropic | Claude training + retrieval |
| PerplexityBot | Perplexity | Answer engine citations |
| Google-Extended | Gemini + AI Overviews training | |
| CCBot | Common Crawl | Open dataset used by many LLMs |
/cart, /checkout, /account, and internal search URLs./wp-admin/ but allow /wp-admin/admin-ajax.php.No. It blocks crawling, not indexing. A blocked URL can still appear in results without a snippet. Use a noindex meta tag to remove pages from the index.
At the domain root: https://yourdomain.com/robots.txt. Subdomains need their own file.
Yes — add a Sitemap: line so crawlers discover your XML sitemap without extra configuration.
I run audits that go beyond robots.txt — crawl budget, indexation, GEO/AI SEO, and growth.
Book a free strategy call