AI crawler access · nature.com
Does nature.com block GPTBot and other AI crawlers?
GPTBot: blockedAs of 2026-10-01, the robots.txt of nature.com blocks 11 of 29 AI crawlers on the home page: GPTBot is blocked, ClaudeBot is blocked and PerplexityBot is blocked.
- Policy: Blocks some AI search (At least one AI search or assistant crawler blocked)
- AI search access: 88% of the AI search and assistant crawlers we track may fetch the home page (these decide whether AI answers can cite the site).
- robots.txt: found at
https://nature.com/robots.txt - llms.txt: no
- Content-Signal: none
- Category: Reference · popularity band: Tranco 101-500
- Last checked: (daily)
All 29 AI crawlers on nature.com
Status for the home page (/) under nature.com's robots.txt. Partially blocked = home page allowed, some paths disallowed.
| Crawler | Operator | robots.txt |
|---|---|---|
| Training crawlers | ||
| GPTBot | OpenAI | blocked |
| ClaudeBot | Anthropic | blocked |
| Google-Extended | blocked | |
| Applebot-Extended | Apple | blocked |
| Meta-ExternalAgent | Meta | blocked |
| Amazonbot | Amazon | blocked |
| MistralAI-Training | Mistral | allowed |
| CCBot | Common Crawl | blocked |
| Bytespider | ByteDance | blocked |
| cohere-ai | Cohere | blocked |
| AI search crawlers | ||
| OAI-SearchBot | OpenAI | allowed |
| Claude-SearchBot | Anthropic | allowed |
| PerplexityBot | Perplexity | blocked |
| Google-CloudVertexBot | allowed | |
| Applebot | Apple | allowed |
| meta-webindexer | Meta | allowed |
| Amzn-SearchBot | Amazon | allowed |
| DuckAssistBot | DuckDuckGo | allowed |
| MistralAI-Index | Mistral | allowed |
| User-triggered fetch crawlers | ||
| ChatGPT-User | OpenAI | blocked |
| Claude-User | Anthropic | allowed |
| Perplexity-User | Perplexity | allowed |
| Google-GeminiNotebook | allowed | |
| Google-Agent | allowed | |
| meta-externalfetcher | Meta | allowed |
| Amzn-User | Amazon | allowed |
| MistralAI-User | Mistral | allowed |
| Ads review crawlers | ||
| OAI-AdsBot | OpenAI | allowed |
| meta-externalads | Meta | allowed |
Compared with other reference sites in the index: 27.9% of them block GPTBot and 23.3% block ClaudeBot (n=43). Across all 853 readable sites: 17% block GPTBot. All reference sites →
History
No change in the latest daily run. We re-check nature.com every day; changes show up here and on the index page.
FAQ
- Does nature.com block GPTBot?
- Yes. As of 2026-10-01, https://nature.com/robots.txt disallows the home page for GPTBot, OpenAI's AI training crawler.
- Does nature.com block ClaudeBot?
- Yes. As of 2026-10-01, https://nature.com/robots.txt disallows the home page for ClaudeBot, Anthropic's AI training crawler.
- Does nature.com block PerplexityBot?
- Yes. As of 2026-10-01, https://nature.com/robots.txt disallows the home page for PerplexityBot, Perplexity's AI search crawler.
- Can ChatGPT search (OAI-SearchBot) crawl nature.com?
- No. As of 2026-10-01, https://nature.com/robots.txt allows OAI-SearchBot, OpenAI's AI search crawler, on the home page.
- Does nature.com block Google-Extended (Gemini training)?
- Yes. As of 2026-10-01, https://nature.com/robots.txt disallows the home page for Google-Extended, Google's AI training crawler.
- Does nature.com have an llms.txt file?
- No. As of 2026-10-01, nature.com does not serve a plain-text /llms.txt file.
- How can I check nature.com or my own site again?
- Use the free AI crawler checker at tools.yukai.uk/free/ai-crawler-checker (no sign-up). To check many sites and get alerts when rules change, use the AI Crawler Access Checker on Apify.
Similar reference sites
- archive.org Open
- w3.org Open
- doi.org No robots.txt
- researchgate.net Open
- wikimedia.org No robots.txt
- who.int Blocks training only
- imdb.com Blocks some AI search
- springer.com Open
- wiley.com Open
- goodreads.com Blocks training only
- arxiv.org Open
- ietf.org Open