AI crawler · Meta · User-triggered fetch
Which websites block meta-externalfetcher?
As of 2026-10-01, 74 of the 853 popular websites whose robots.txt we could read (8.7%) block meta-externalfetcher, Meta's user-triggered AI fetcher, on the home page; 11 more block it on some paths. Among the top 100 sites: 14.5%.
8.7%
of 853 readable sites block meta-externalfetcher
14.5%
of the top 100
#13
most blocked of 29 AI crawlers
768
sites allow it on the home page
Blocked by category
| Category | Sites read | Block meta-externalfetcher |
|---|---|---|
| News & media | 126 | 27% |
| Social | 48 | 18.8% |
| Other | 260 | 7.7% |
| Travel | 13 | 7.7% |
| Reference | 43 | 7% |
| E-commerce | 58 | 5.2% |
| Entertainment | 68 | 4.4% |
| Tech | 154 | 0.6% |
| AI | 9 | 0% |
| Education | 20 | 0% |
| Finance | 16 | 0% |
| Government | 33 | 0% |
| Telecom | 5 | 0% |
Popular sites that block meta-externalfetcher (74)
- facebook.com Social
- instagram.com Social
- twitter.com Social
- linkedin.com Social
- amazon.com E-commerce
- netflix.com Entertainment
- x.com Social
- reddit.com Social
- amazonvideo.com Entertainment
- snapchat.com Social
- nytimes.com News & media
- flickr.com Other
- meraki.com Tech
- imdb.com Reference
- launchpad.net Other
- reuters.com News & media
- weibo.com Social
- trustpilot.com Other
- bloomberg.com News & media
- wsj.com News & media
- ft.com News & media
- wired.com News & media
- uol.com.br News & media
- usatoday.com News & media
- threads.com Social
- telegraph.co.uk News & media
- primevideo.com Entertainment
- dailymail.com News & media
- fb.com Other
- cnet.com News & media
- people.com News & media
- academia.edu Reference
- theverge.com News & media
- repubblica.it News & media
- investopedia.com News & media
- flashscore.com News & media
- theatlantic.com News & media
- bfmtv.com News & media
- huffpost.com News & media
- nypost.com News & media
- theconversation.com News & media
- ad.nl News & media
- nu.nl News & media
- elconfidencial.com News & media
- amap.com Travel
- hln.be News & media
- emag.ro E-commerce
- themoviedb.org Reference
- actu.fr News & media
- huffingtonpost.com News & media
- dw.com News & media
- iltalehti.fi News & media
- ekstrabladet.dk News & media
- tjk.org Other
- buzzfeed.com Other
- aka.ms Other
- sportfilm800.com News & media
- twkan.com Other
- rocketmoney.dev Other
- qcloud.com Other
- newyorker.com Other
- blackhub.team Other
- arstechnica.com News & media
- zdnet.com News & media
- adrta.com Other
- marketwatch.com E-commerce
- merkur.de Other
- wikihow.com Other
- lnkd.in Other
- sogou.com Other
- ign.com Other
- salesforceliveagent.com Other
- sedoparking.com Other
- washingtonpost.com News & media
Sites that block meta-externalfetcher on some paths (11)
- pinterest.com Social
- hostinger.com Tech
- indeed.com Other
- sohu.com Tech
- statista.com News & media
- capcut.com Other
- jotform.com Tech
- disneyplus.com Entertainment
- jimdo.com Tech
- deezer.com Other
- hulu.com Entertainment
Trend
The daily series for meta-externalfetcher starts on 2026-10-01; a trend appears here as days accumulate.
Latest changes
- dailymotion.com: blocked → allowed
Block or allow meta-externalfetcher
User-agent: meta-externalfetcher Disallow: /
Replace Disallow: / with Allow: / to let it in. A site that blocks training crawlers but wants to be cited in AI answers usually keeps the search and user-triggered crawlers allowed.
FAQ
- How many websites block meta-externalfetcher?
- As of 2026-10-01, 74 of the 853 popular websites whose robots.txt we could read (8.7%) block meta-externalfetcher, Meta's user-triggered AI fetcher, on the home page; 11 more block it on some paths. Among the top 100 sites: 14.5%.
- What is meta-externalfetcher?
- meta-externalfetcher is Meta's user-triggered AI fetcher: it fetches a page when a user asks the assistant about it. It is the #13 most blocked of the 29 AI crawlers in the index.
- How do I block meta-externalfetcher in robots.txt?
- Add these lines to https://<your-site>/robots.txt: "User-agent: meta-externalfetcher" followed by "Disallow: /". To allow it, use "Allow: /" instead. Well-behaved crawlers read robots.txt before fetching pages; firewall rules are needed to stop crawlers that ignore it.
- How do I check whether my site blocks meta-externalfetcher?
- Use the free AI crawler checker at tools.yukai.uk/free/ai-crawler-checker: it shows the status of meta-externalfetcher and 28 other AI crawlers with the exact robots.txt rule.
Other AI crawlers
Also from Meta: Meta-ExternalAgent (14.4%), meta-webindexer (6.3%), meta-externalads (3.8%).
- CCBot 19.5%
- Bytespider 17.7%
- GPTBot 17%
- ClaudeBot 16.3%
- Meta-ExternalAgent 14.4%
- Applebot-Extended 14.3%
- Google-Extended 14%
- cohere-ai 13.7%
- Amazonbot 13.1%
- PerplexityBot 12.7%
- ChatGPT-User 9.6%
- Perplexity-User 8.9%
- MistralAI-User 8.4%
- Claude-User 8.3%
- Claude-SearchBot 8.1%
- DuckAssistBot 8.1%
- Google-CloudVertexBot 7.2%
- OAI-SearchBot 6.7%
- meta-webindexer 6.3%
- Amzn-SearchBot 5.9%
- Amzn-User 5.5%
- MistralAI-Index 4.7%
- MistralAI-Training 4.7%
- Google-Agent 4.5%
- OAI-AdsBot 4.3%
- Google-GeminiNotebook 4.3%
- meta-externalads 3.8%
- Applebot 2.6%