# Google-Extended vs OAI-SearchBot

Google-Extended and OAI-SearchBot compared on 12 documented fields, from 18 claims, last verified 2026-09-15.

| Field | Google-Extended | OAI-SearchBot | Comparison |
| --- | --- | --- | --- |
| cites_sources | not documented | yes | different |
| honours_crawl_delay | n/a | not_documented | different |
| observed_arrival_before_sitemap | not documented | yes | different |
| publishes_ip_list | no | yes | different |
| purpose_documented | policy_only | search | different |
| reads_llms_txt | not documented | not_documented | different |
| respects_robots_txt | n/a | yes | different |
| robots_token_differs_from_ua | yes | not documented | different |
| separate_search_and_training_tokens | yes | yes | same |
| user_agent_string_full | not documented | Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.4; +https://openai.com/searchbot | different |
| verifiable_by_rdns | n/a | not_documented | different |
| what_it_controls | Gemini model training and grounding, not Search inclusion | not documented | different |

## Differences

| Field | Google-Extended statement | OAI-SearchBot statement |
| --- | --- | --- |
| cites_sources | not documented | OAI-SearchBot surfaces crawled websites as links in ChatGPT search answers. |
| honours_crawl_delay | Google does not publish any Crawl-delay behaviour for the Google-Extended token. | OpenAI does not publish whether OAI-SearchBot honours the robots.txt Crawl-delay directive. |
| observed_arrival_before_sitemap | not documented | OAI-SearchBot first requested Rattlesnakes By Mail at 2026-09-13T21:41:14Z, before the Rattlesnakes By Mail sitemap was submitted to Bing Webmaster Tools shortly before 2026-09-15T08:38Z and to Google Search Console at 2026-09-15T09:00Z. OAI-SearchBot made 3 requests to Rattlesnakes By Mail between 2026-09-13T20:00Z and 2026-09-15T09:00Z, all 3 to /robots.txt. |
| publishes_ip_list | Google publishes no IP list for Google-Extended because Google-Extended has no separate HTTP request user agent string. | OpenAI publishes the IP address ranges used by OAI-SearchBot as a JSON file at https://openai.com/searchbot.json. |
| purpose_documented | Google-Extended is a standalone product token that controls whether Google uses crawled content to train Gemini models. | OAI-SearchBot surfaces websites in search results in ChatGPT search features. |
| reads_llms_txt | not documented | OpenAI does not publish whether OAI-SearchBot reads llms.txt files. |
| respects_robots_txt | Google-Extended is a robots.txt token that publishers set to control use of crawled content for Gemini training. | OAI-SearchBot honours robots.txt opt-outs set for the OAI-SearchBot token. |
| robots_token_differs_from_ua | Google-Extended is a robots.txt control token and does not correspond to a crawler user agent string. | not documented |
| user_agent_string_full | not documented | OAI-SearchBot sends the user agent string Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.4; +https://openai.com/searchbot. |
| verifiable_by_rdns | Google does not publish a reverse-DNS verification method for Google-Extended. | OpenAI does not publish a reverse-DNS verification method for OAI-SearchBot. |
| what_it_controls | Google-Extended controls AI training and grounding in Google systems and does not control Google Search inclusion. | not documented |

## Sources

- https://developers.google.com/search/docs/appearance/ai-features
- https://developers.google.com/search/docs/crawling-indexing/google-common-crawlers
- https://developers.openai.com/api/docs/bots
- https://openai.com/searchbot.json
- https://rattlesnakesbymail.com/observed/oai-searchbot
