OAI-SearchBot
Search and retrieval for OpenAI search experiences.
Operator: OpenAI
Category: AI Search
User-agent: OAI-SearchBot
Robots.txt: Scored
Reference · Official IP ranges
Updated 2026-08-03
Official IP-range links appear when the operator publishes them. Match both the user agent and current IP ranges when reviewing WAF rules or request logs.
Search and retrieval for OpenAI search experiences.
Operator: OpenAI
Category: AI Search
User-agent: OAI-SearchBot
Robots.txt: Scored
Reference · Official IP ranges
Required crawler for ChatGPT Ads landing-page validation and review; not used for foundation-model training.
Operator: OpenAI
Category: Ad Validation
User-agent: OAI-AdsBot
Robots.txt: Scored
Reference · Official IP ranges
User-triggered visits from ChatGPT and Custom GPT actions; not an automatic web crawler.
Operator: OpenAI
Category: User Triggered
User-agent: ChatGPT-User
Robots.txt: Not scored: may not apply
Reference · Official IP ranges
OpenAI web crawler used for model improvement according to OpenAI documentation.
Operator: OpenAI
Category: AI Training
User-agent: GPTBot
Robots.txt: Scored
Reference · Official IP ranges
Search and retrieval crawler for Claude experiences.
Operator: Anthropic
Category: AI Search
User-agent: Claude-SearchBot
Robots.txt: Scored
Reference
User-triggered fetches from Claude that site owners can control through robots.txt.
Operator: Anthropic
Category: User Triggered
User-agent: Claude-User
Robots.txt: Scored
Reference
Anthropic crawler for model-related web access.
Operator: Anthropic
Category: AI Training
User-agent: ClaudeBot
Robots.txt: Scored
Reference
Perplexity crawler for search and answer experiences.
Operator: Perplexity
Category: AI Search
User-agent: PerplexityBot
Robots.txt: Scored
Reference · Official IP ranges
User-triggered fetches used when Perplexity answers a user request.
Operator: Perplexity
Category: User Triggered
User-agent: Perplexity-User
Robots.txt: Not scored: generally ignored
Reference · Official IP ranges
Google product improvement control token for Gemini and Vertex AI training use.
Operator: Google
Category: AI Training
User-agent: Google-Extended
Robots.txt: Scored
Reference
Google crawler token for crawls requested by site owners when building Vertex AI Agents; not a Google Search ranking signal.
Operator: Google
Category: Site-owner Requested
User-agent: Google-CloudVertexBot
Robots.txt: Scored
Reference
Classic Google Search crawling and indexing.
Operator: Google
Category: Classic Search
User-agent: Googlebot
Robots.txt: Scored
Reference
Classic Bing Search crawling and indexing.
Operator: Microsoft
Category: Classic Search
User-agent: bingbot
Robots.txt: Scored
Reference
Common Crawl web archive crawler used by many downstream projects.
Operator: Common Crawl
Category: Common AI
User-agent: CCBot
Robots.txt: Scored
Reference
Apple web crawler for search-style experiences such as Siri and Spotlight; Applebot-Extended controls AI-related use separately.
Operator: Apple
Category: Classic Search
User-agent: Applebot
Robots.txt: Scored
Reference
Apple extension token for AI-related use controls.
Operator: Apple
Category: AI Training
User-agent: Applebot-Extended
Robots.txt: Scored
Reference
Meta crawler token commonly used for AI training controls.
Operator: Meta
Category: AI Training
User-agent: Meta-ExternalAgent
Robots.txt: Scored
Reference
ByteDance crawler often discussed in AI crawler policies.
Operator: ByteDance
Category: AI Training
User-agent: Bytespider
Robots.txt: Scored
Reference