Example result
Here's what you get — a real example generated by the tool:
## robots.txt
```
User-agent: GPTBot
Disallow: /
User-agent: ChatGPT-User
Disallow: /
User-agent: ClaudeBot
Disallow: /
User-agent: Claude-User
Disallow: /
User-agent: Google-Extended
Disallow: /
User-agent: CCBot
Disallow: /
User-agent: Bytespider
Disallow: /
User-agent: PerplexityBot
Disallow: /
User-agent: Amazonbot
Disallow: /
User-agent: Applebot-Extended
Disallow: /
```
## llms.txt
```
# myblog.com
Personal blog about remote work and productivity.
## AI Training
Disallow: GPTBot, ChatGPT-User, ClaudeBot, Claude-User, Google-Extended, CCBot, Bytespider, PerplexityBot, Amazonbot, Applebot-Extended
```
## What Each Bot Does
- **OpenAI — GPTBot / ChatGPT-User:** Scrapes pages to train GPT models and to fetch live pages when ChatGPT browses the web.
- **Anthropic — ClaudeBot / Claude-User:** Crawls content to train Claude models and to answer live user queries inside Claude.
- **Google — Google-Extended:** Controls whether your content can be used to train Gemini and other Google AI models (separate from normal Search indexing).
- **Common Crawl — CCBot:** A public web archive many AI labs download and train on, so blocking it cuts off several training pipelines at once.
- **ByteDance — Bytespider:** TikTok's parent company's crawler, used to gather training data for its AI products.
- **Perplexity — PerplexityBot:** Powers Perplexity's AI answer engine, which can reproduce your content directly in chat answers instead of sending you a visitor.
- **Amazon — Amazonbot:** Collects data for Amazon's AI and Alexa-related products.
- **Apple — Applebot-Extended:** Governs use of your content for training Apple Intelligence features.
## Important Notes
- robots.txt is a voluntary standard — reputable AI companies honor it, but it cannot force a bot to stay away.
- This does not affect normal search visibility: Googlebot and Bingbot are untouched unless you add them yourself.
- Upload robots.txt to your site's root (https://myblog.com/robots.txt) and llms.txt the same way (https://myblog.com/llms.txt).
Frequently Asked Questions
What file do I get?
A .docx file containing a ready-to-paste robots.txt block, an optional llms.txt block, and a plain-English explainer of what each AI bot does with your content — unlike ChatGPT, which can't track which AI crawlers currently exist.
Why pay €9.99 when robots.txt is free to write myself?
Writing it yourself means researching which of 15+ AI companies run crawlers, finding their exact current user-agent strings, and getting the syntax right. This tool does that research for you and hands you the finished file in 60 seconds.
How do you know the AI bot list is current?
The tool is built on an actively maintained reference of AI crawler user-agents from OpenAI, Google, Anthropic, Common Crawl, Perplexity, ByteDance, and other major AI companies, updated as new crawlers launch.
Will this affect my Google Search ranking?
No. Blocking AI training crawlers like Google-Extended, GPTBot, or ClaudeBot does not affect normal search indexing bots like Googlebot or Bingbot unless you explicitly add them to the block list.
How fast do I get my file?
Within 60 seconds of payment. You'll see the generated robots.txt and llms.txt content on screen and can download the .docx file immediately.
Is robots.txt legally binding — will AI companies actually respect it?
robots.txt is a voluntary standard. Major AI companies like OpenAI, Anthropic, and Google publicly state they honor it, but it can't physically stop a bot that chooses to ignore it. It remains the standard first line of defense and is required reading for any compliant crawler.
Can I get a refund?
Yes. If the generated files don't work for your site, contact us within 24 hours for a full refund, no questions asked.