AI Crawler Index · Common Crawl · CCBot
Which .es (Spain) websites block CCBot?
As of 2026-10-03, 8 of the 36 .es (Spain) websites whose robots.txt we could read (22.2%) block CCBot, Common Crawl's AI training crawler, on the home page, compared with 12.8% of all sites.
22.2%
of 36 .es (Spain) websites block CCBot
12.8%
of all sites
8
sites block it
All 8 .es (Spain) websites that block CCBot
| # | Site | Policy | Crawlers blocked |
|---|---|---|---|
| 442 | elmundo.es | Blocks some AI search | 6 of 29 |
| 2206 | aepd.es | Blocks all AI search | 29 of 29 |
| 2433 | 20minutos.es | Blocks training only | 8 of 29 |
| 3029 | larazon.es | Blocks training only | 6 of 29 |
| 4021 | heraldo.es | Blocks training only | 8 of 29 |
| 6169 | tvguia.es | Blocks some AI search | 7 of 29 |
| 7030 | cope.es | Blocks some AI search | 10 of 29 |
| 9816 | cervantes.es | Blocks some AI search | 28 of 29 |
Popular .es (Spain) websites that allow CCBot
Other AI crawlers and .es (Spain) websites
- GPTBot 25%
- ClaudeBot 19.4%
- PerplexityBot 13.9%
- Google-Extended 8.3%
- Bytespider 19.4%
- Applebot-Extended 13.9%
- OAI-SearchBot 13.9%
CCBot in other segments
- news websites 35.8%
- .fr (France) websites 35.3%
- .it (Italy) websites 32.7%
- social websites 30.8%
- .au (Australia) websites 29.7%
- reference websites 28.9%
- .nl (Netherlands) websites 27.1%
- .uk (United Kingdom) websites 24.4%
- .de (Germany) websites 23.6%
- .pl (Poland) websites 20.8%
- .app websites 16.7%
- .ch (Switzerland) websites 16.7%
- .com websites 14%
- .net websites 13.9%
FAQ
- Which .es (Spain) websites block CCBot?
- elmundo.es, aepd.es, 20minutos.es, larazon.es, heraldo.es, tvguia.es, cope.es, cervantes.es.
- How many .es (Spain) websites block CCBot?
- As of 2026-10-03, 8 of the 36 .es (Spain) websites whose robots.txt we could read (22.2%) block CCBot, Common Crawl's AI training crawler, on the home page, compared with 12.8% of all sites.
- How do I block or allow CCBot on my site?
- Add "User-agent: CCBot" followed by "Disallow: /" to https://<your-site>/robots.txt to block it, or "Allow: /" to allow it. Check the result with the free AI crawler checker at tools.yukai.uk/free/ai-crawler-checker.