// Interactive Policy Builder · Free

AI robots.txt Generator

Configure fine-grained crawler permissions for 12+ AI search and training bots. Allow citation traffic on ChatGPT & Perplexity while controlling unauthorized LLM model scrapers.

// Free Developer Tools: 🤖AI Robots.txt 📄llms.txt Generator 🔍GPTBot Checker ⚡Live Crawlability 📊28-Pt GEO Score
Strategy Presets:

1. Global & General Settings

Universal Crawlers

2. AI Crawler Permission Matrix

Toggle permission for each major AI engine individually.

12 Crawlers
// robots.txt Preview

        
Place at: /robots.txt Test site live →

AI Crawlers Reference & Robots.txt Cheat-Sheet

Standardized syntax and recommendations for the major search engines and foundation model crawlers.

OAI-SearchBot Citations

Powers real-time search results and source links in ChatGPT Search. Recommended: Always Allow.

User-agent: OAI-SearchBot
Allow: /
PerplexityBot Citations

Indexes content for Perplexity AI answers and footnotes. Recommended: Always Allow.

User-agent: PerplexityBot
Allow: /
GPTBot Training

Collects training data for OpenAI models. Disallowing does not affect ChatGPT Search citations.

User-agent: GPTBot
Disallow: /
ClaudeBot Training

Anthropic model training crawler. Use Claude-Web to control interactive browsing in Claude chat.

User-agent: ClaudeBot
Disallow: /
Google-Extended Training

Controls Gemini and Vertex AI training. Disallowing does not affect Googlebot Search indexation.

User-agent: Google-Extended
Disallow: /
Applebot-Extended AI / Siri

Controls content ingestion for Apple Intelligence and Siri generative search features.

User-agent: Applebot-Extended
Allow: /

Is Your robots.txt Working on Your Live Website?

Robots directives can be easily overridden by CDN caching, Cloudflare WAF, or server header misconfigurations.