| cites_sources | not documented | Anthropic does not publish whether Claude search responses cite pages crawled by Claude-SearchBot. |
| honours_crawl_delay | Apple does not publish a Crawl-delay behaviour for Applebot-Extended. | Claude-SearchBot respects the Crawl-delay directive in robots.txt when crawling domains. |
| publishes_ip_list | Apple publishes no IP list specific to Applebot-Extended and documents one shared crawl identity for Applebot. | Anthropic publishes the IP address ranges used by Claude-SearchBot as a JSON file at https://claude.com/crawling/bots.json. |
| purpose_documented | Applebot-Extended controls whether website content is used to train Apple general purpose foundation models. | Claude-SearchBot analyses online content to improve the relevance and accuracy of Claude search responses. |
| reads_llms_txt | not documented | Anthropic does not publish whether Claude-SearchBot reads llms.txt files. |
| respects_robots_txt | Applebot-Extended is a robots.txt token that publishers disallow to opt out of generative model training. | Claude-SearchBot honours industry standard robots.txt directives that signal do not crawl. |
| robots_token_differs_from_ua | Applebot-Extended is a robots.txt control token and does not correspond to a separate crawler user agent string. | not documented |
| user_agent_string_full | not documented | Anthropic does not publish a full user agent string for Claude-SearchBot. |
| verifiable_by_rdns | Apple does not publish a reverse-DNS verification method specific to Applebot-Extended. | Anthropic does not publish a reverse-DNS verification method for Claude-SearchBot. |
| what_it_controls | Applebot-Extended controls the AI training use of crawled content and does not control search indexing. | not documented |