Back to GEO LibraryTECHNICAL / CRAWLER POLICY

AI crawler controls by platform

AI access policy should be explicit by crawler and purpose. A blanket allow or block can create consequences the publisher did not intend.

Updated
20/08/2026
Publisher
NobleJackal
Version
1.0
Source review
Primary sources reviewed
01

Google uses the Search foundation

Google states that AI features rely on existing Search systems. There is no special AI schema requirement; index eligibility, snippets and normal Search controls remain relevant.

02

OpenAI separates search from training

OAI-SearchBot supports search discovery, while GPTBot is a separate control associated with potential model training. Publishers should decide on each purpose independently.

03

Anthropic and Perplexity publish distinct controls

Anthropic and Perplexity document crawler identities and retrieval behaviour. Robots rules should be paired with live CDN or WAF checks because a policy file alone cannot prove successful access.

Practical checklist

  • Document the purpose of each crawler rule
  • Separate search retrieval from training preferences
  • Test live responses with the published user agents
  • Review official documentation after platform changes

Official sources

  1. Google Search CentralAI features and your website
  2. OpenAIOverview of OpenAI crawlers
  3. Anthropic Help CenterAnthropic web crawlers
  4. Perplexity DocsPerplexity crawlers