Start Tracking Your Brand
Take Actions
Agent Analytics
Crawlability helps you understand which AI crawlers are allowed or blocked from accessing your website based on your robots.txt configuration.
Simply select one of your tracked domains, and Cite AI automatically analyzes your robots.txt file against supported AI crawlers. No additional setup is required.
This helps you identify whether important AI platforms can crawl your content and whether any robots.txt rules are unintentionally limiting your AI visibility.
Cite AI evaluates your robots.txt rules for every supported AI crawler and displays their crawl permissions.
For each crawler you’ll see:
| Field | Description |
| Bot | The crawler’s user-agent (for example, GPTBot or ClaudeBot). |
| Platform | The AI provider behind the crawler, such as OpenAI, Google, Anthropic, or Perplexity. |
| Bot Type | The crawler’s primary purpose, including Training, Search, User Retrieval, or Other. |
| Status | Whether the crawler is Allowed, Partially Allowed, or Blocked. |
| Reason | Which robots.txt rule determined the result, including explicit bot rules or inherited wildcard rules. |
Use the available filters to search by platform, bot type, or crawl status.
The URL Tester allows you to verify crawl permissions for any page on your website.
Simply enter a URL from your domain to instantly see which AI crawlers can access it and which are blocked.
This is useful for:
If an AI crawler is blocked by your robots.txt file, it cannot access the content on that page.
Blocked content may be less likely to:
Use Crawlability to:
Cite AI groups supported crawlers based on their primary purpose.
Training bots collect publicly available content that may be used to train or improve AI models.
Examples include:
Search bots retrieve live web content to answer user queries and generate up-to-date responses.
Examples include:
These bots visit websites while responding to a specific user’s request inside an AI assistant.
Examples include:
Some supported crawlers perform specialized functions that don’t fall into the categories above, such as research, indexing, metadata collection, or platform-specific services.
These bots are grouped under Other and are monitored separately within Crawlability.
By understanding which AI crawlers can access your website, you can ensure your content remains discoverable across the AI platforms that matter most while maintaining full control over your robots.txt configuration.