Rattlesnakes By Mail

Google-Extended vs Googlebot

Google-Extended and Googlebot compared on 12 documented fields, from 18 claims, last verified 2026-09-15.

FieldGoogle-ExtendedGooglebotComparison
cites_sourcesnot documentedyesdifferent
honours_crawl_delayn/anodifferent
observed_verified_request_ratio_censusnot documented0.991different
publishes_ip_listnoyesdifferent
purpose_documentedpolicy_onlysearchdifferent
reads_llms_txtnot documentednodifferent
respects_robots_txtn/ayesdifferent
robots_token_differs_from_uayesnot documenteddifferent
separate_search_and_training_tokensyesyessame
user_agent_string_fullnot documentedMozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Googlebot/2.1; +http://www.google.com/bot.html)different
verifiable_by_rdnsn/ayesdifferent
what_it_controlsGemini model training and grounding, not Search inclusionnot documenteddifferent

Differences

FieldGoogle-Extended statementGooglebot statement
cites_sourcesnot documentedGoogle AI Overviews and AI Mode return AI generated responses with links to supporting websites.
honours_crawl_delayGoogle does not publish any Crawl-delay behaviour for the Google-Extended token.Google does not support the robots.txt Crawl-delay directive for Googlebot.
observed_verified_request_ratio_censusnot documentedGooglebot requests to Rattlesnakes By Mail matched Google's published IP ranges on 108 of 109 requests between 2026-09-13T20:00Z and 2026-09-15T09:00Z. Googlebot's first request to Rattlesnakes By Mail in that window was /robots.txt at 2026-09-14T11:53:00Z.
publishes_ip_listGoogle publishes no IP list for Google-Extended because Google-Extended has no separate HTTP request user agent string.Google publishes the IP address ranges used by Googlebot at https://developers.google.com/static/crawling/ipranges/common-crawlers.json.
purpose_documentedGoogle-Extended is a standalone product token that controls whether Google uses crawled content to train Gemini models.Googlebot crawls pages for Google Search, Google Images, Google Video, Google News, and Discover.
reads_llms_txtnot documentedGoogle Search does not use llms.txt files.
respects_robots_txtGoogle-Extended is a robots.txt token that publishers set to control use of crawled content for Gemini training.Googlebot obeys robots.txt rules when crawling automatically.
robots_token_differs_from_uaGoogle-Extended is a robots.txt control token and does not correspond to a crawler user agent string.not documented
user_agent_string_fullnot documentedGooglebot sends the user agent string Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Googlebot/2.1; +http://www.google.com/bot.html).
verifiable_by_rdnsGoogle does not publish a reverse-DNS verification method for Google-Extended.Googlebot requests are verifiable by reverse DNS lookup resolving to googlebot.com, google.com, or googleusercontent.com.
what_it_controlsGoogle-Extended controls AI training and grounding in Google systems and does not control Google Search inclusion.not documented

Sources