# Claude-SearchBot vs Google-Extended

Claude-SearchBot and Google-Extended compared on 11 documented fields, from 17 claims, last verified 2026-09-13.

| Field | Claude-SearchBot | Google-Extended | Comparison |
| --- | --- | --- | --- |
| cites_sources | not_documented | not documented | different |
| honours_crawl_delay | yes | n/a | different |
| publishes_ip_list | yes | no | different |
| purpose_documented | search | policy_only | different |
| reads_llms_txt | not_documented | not documented | different |
| respects_robots_txt | yes | n/a | different |
| robots_token_differs_from_ua | not documented | yes | different |
| separate_search_and_training_tokens | yes | yes | same |
| user_agent_string_full | not_documented | not documented | different |
| verifiable_by_rdns | not_documented | n/a | different |
| what_it_controls | not documented | Gemini model training and grounding, not Search inclusion | different |

## Differences

| Field | Claude-SearchBot statement | Google-Extended statement |
| --- | --- | --- |
| cites_sources | Anthropic does not publish whether Claude search responses cite pages crawled by Claude-SearchBot. | not documented |
| honours_crawl_delay | Claude-SearchBot respects the Crawl-delay directive in robots.txt when crawling domains. | Google does not publish any Crawl-delay behaviour for the Google-Extended token. |
| publishes_ip_list | Anthropic publishes the IP address ranges used by Claude-SearchBot as a JSON file at https://claude.com/crawling/bots.json. | Google publishes no IP list for Google-Extended because Google-Extended has no separate HTTP request user agent string. |
| purpose_documented | Claude-SearchBot analyses online content to improve the relevance and accuracy of Claude search responses. | Google-Extended is a standalone product token that controls whether Google uses crawled content to train Gemini models. |
| reads_llms_txt | Anthropic does not publish whether Claude-SearchBot reads llms.txt files. | not documented |
| respects_robots_txt | Claude-SearchBot honours industry standard robots.txt directives that signal do not crawl. | Google-Extended is a robots.txt token that publishers set to control use of crawled content for Gemini training. |
| robots_token_differs_from_ua | not documented | Google-Extended is a robots.txt control token and does not correspond to a crawler user agent string. |
| user_agent_string_full | Anthropic does not publish a full user agent string for Claude-SearchBot. | not documented |
| verifiable_by_rdns | Anthropic does not publish a reverse-DNS verification method for Claude-SearchBot. | Google does not publish a reverse-DNS verification method for Google-Extended. |
| what_it_controls | not documented | Google-Extended controls AI training and grounding in Google systems and does not control Google Search inclusion. |

## Sources

- https://claude.com/crawling/bots.json
- https://developers.google.com/search/docs/appearance/ai-features
- https://developers.google.com/search/docs/crawling-indexing/google-common-crawlers
- https://support.claude.com/en/articles/8896518-does-anthropic-crawl-content-from-the-web-and-how-can-site-owners-block-the-crawler
