TidyToolsAI Crawler Index · data as of · updated daily

AI crawler access · phys.org

Does phys.org block GPTBot and other AI crawlers?

GPTBot: partially blockedAs of 2026-10-03, the robots.txt of phys.org blocks 0 of 29 AI crawlers on the home page and partially blocks 7 more: GPTBot is partially blocked, ClaudeBot is partially blocked and PerplexityBot is allowed.

All 29 AI crawlers on phys.org

Status for the home page (/) under phys.org's robots.txt. Partially blocked = home page allowed, some paths disallowed.

CrawlerOperatorrobots.txt
Training crawlers
GPTBotOpenAIpartial
ClaudeBotAnthropicpartial
Google-ExtendedGooglepartial
Applebot-ExtendedApplepartial
Meta-ExternalAgentMetaallowed
AmazonbotAmazonallowed
MistralAI-TrainingMistralallowed
CCBotCommon Crawlpartial
BytespiderByteDancepartial
cohere-aiCoherepartial
AI search crawlers
OAI-SearchBotOpenAIallowed
Claude-SearchBotAnthropicallowed
PerplexityBotPerplexityallowed
Google-CloudVertexBotGoogleallowed
ApplebotAppleallowed
meta-webindexerMetaallowed
Amzn-SearchBotAmazonallowed
DuckAssistBotDuckDuckGoallowed
MistralAI-IndexMistralallowed
User-triggered fetch crawlers
ChatGPT-UserOpenAIallowed
Claude-UserAnthropicallowed
Perplexity-UserPerplexityallowed
Google-GeminiNotebookGoogleallowed
Google-AgentGoogleallowed
meta-externalfetcherMetaallowed
Amzn-UserAmazonallowed
MistralAI-UserMistralallowed
Ads review crawlers
OAI-AdsBotOpenAIallowed
meta-externaladsMetaallowed

Compared with other other sites in the index: 11.6% of them block GPTBot and 10.4% block ClaudeBot (n=6685). Across all 8,357 readable sites: 12% block GPTBot. All other sites →

On .org domains: 11.1% of 514 readable sites block GPTBot and 8.6% serve an llms.txt. All .org sites →

History

No change recorded since daily tracking started on 2026-10-02. We re-check phys.org every day; a change in its AI crawler rules shows up here and on the index page.

FAQ

Does phys.org block GPTBot?
Partly. As of 2026-10-03, https://phys.org/robots.txt allows GPTBot (OpenAI's AI training crawler) on the home page but disallows some paths.
Does phys.org block ClaudeBot?
Partly. As of 2026-10-03, https://phys.org/robots.txt allows ClaudeBot (Anthropic's AI training crawler) on the home page but disallows some paths.
Does phys.org block PerplexityBot?
No. As of 2026-10-03, https://phys.org/robots.txt allows PerplexityBot, Perplexity's AI search crawler, on the home page.
Can ChatGPT search (OAI-SearchBot) crawl phys.org?
No. As of 2026-10-03, https://phys.org/robots.txt allows OAI-SearchBot, OpenAI's AI search crawler, on the home page.
Does phys.org block Google-Extended (Gemini training)?
Partly. As of 2026-10-03, https://phys.org/robots.txt allows Google-Extended (Google's AI training crawler) on the home page but disallows some paths.
Does phys.org have an llms.txt file?
Yes. phys.org serves a plain-text /llms.txt file (checked 2026-10-03).
How can I check phys.org or my own site again?
Use the free AI crawler checker at tools.yukai.uk/free/ai-crawler-checker (no sign-up). To check many sites and get alerts when rules change, use the AI Crawler Access Checker on Apify.

Similar .org sites

All sites with an llms.txt · Which .org sites block GPTBot · What is phys.org built with? (Tech Stack Index)

All 10,000 sites · All 29 AI crawlers · AI Crawler Index