Rattlesnakes By Mail

ClaudeBot vs Perplexity-User

ClaudeBot and Perplexity-User compared on 14 documented fields, from 21 claims, last verified 2026-09-15.

FieldClaudeBotPerplexity-UserComparison
anti_circumvention_stancedocumentednot documenteddifferent
cites_sourcesnot documentedyesdifferent
honours_crawl_delayyesnot_documenteddifferent
observed_fetches_jsonyesnot documenteddifferent
observed_fetches_markdownyesnot documenteddifferent
observed_first_request_path/sitemap.xmlnot documenteddifferent
observed_format_pass_orderhtml_json_mdnot documenteddifferent
publishes_ip_listyesyessame
purpose_documentedtraininguser_fetchdifferent
reads_llms_txtnot documentednot_documenteddifferent
respects_robots_txtyesnodifferent
separate_search_and_training_tokensyesnot_documenteddifferent
user_agent_string_fullnot_documentedMozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Perplexity-User/1.0; +https://perplexity.ai/perplexity-user)different
verifiable_by_rdnsnot_documentednot_documentedsame

Differences

FieldClaudeBot statementPerplexity-User statement
anti_circumvention_stanceClaudeBot respects anti-circumvention technologies and does not bypass CAPTCHAs on crawled sites.not documented
cites_sourcesnot documentedPerplexity includes a link to each page that Perplexity-User visits in the Perplexity response.
honours_crawl_delayAnthropic supports the non-standard robots.txt Crawl-delay extension for ClaudeBot.Perplexity does not publish whether Perplexity-User honours the robots.txt Crawl-delay directive.
observed_fetches_jsonClaudeBot made 177 JSON requests to Rattlesnakes By Mail between 2026-09-13T20:00Z and 2026-09-15T09:00Z, covering the .json twin of 170 base paths. The .json URLs of Rattlesnakes By Mail appear in no entry of the 171 URL sitemap.not documented
observed_fetches_markdownClaudeBot fetched the .md twin of 170 base paths on Rattlesnakes By Mail between 2026-09-13T20:00Z and 2026-09-15T09:00Z, 170 Markdown requests of 542 ClaudeBot requests in that window. The .md URLs of Rattlesnakes By Mail appear in no entry of the 171 URL sitemap.not documented
observed_first_request_pathClaudeBot's first request to Rattlesnakes By Mail between 2026-09-13T20:00Z and 2026-09-15T09:00Z was /sitemap.xml at 2026-09-14T19:53:45Z. ClaudeBot fetched /robots.txt on the second request to Rattlesnakes By Mail.not documented
observed_format_pass_orderClaudeBot fetched Rattlesnakes By Mail in three format passes between 2026-09-13T20:00Z and 2026-09-15T09:00Z, HTML from 2026-09-14T20:52Z to 2026-09-14T20:57Z, then JSON from 2026-09-14T21:37Z to 2026-09-14T21:56Z, then Markdown from 2026-09-14T21:37Z to 2026-09-14T21:57Z. The same order holds on 170 of 170 base paths that ClaudeBot fetched in all three formats on Rattlesnakes By Mail.not documented
purpose_documentedClaudeBot collects web content for training Anthropic generative AI models.Perplexity-User visits web pages to answer questions that users ask Perplexity.
reads_llms_txtnot documentedPerplexity does not publish whether Perplexity-User reads llms.txt files.
respects_robots_txtClaudeBot honours industry standard robots.txt directives that signal do not crawl.Perplexity-User ignores robots.txt rules because a user requests each Perplexity-User fetch.
separate_search_and_training_tokensAnthropic separates training and search controls into two robots.txt tokens: ClaudeBot for training and Claude-SearchBot for search.Perplexity does not document Perplexity-User as a search or training opt-out token.
user_agent_string_fullAnthropic does not publish a full user agent string for ClaudeBot.Perplexity-User sends the user agent string Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Perplexity-User/1.0; +https://perplexity.ai/perplexity-user).

Sources