AI Crawler Index · OpenAI · GPTBot
Which government websites block GPTBot?
As of 2026-10-03, 8 of the 210 government websites whose robots.txt we could read (3.8%) block GPTBot, OpenAI's AI training crawler, on the home page, compared with 12% of all sites.
3.8%
of 210 government websites block GPTBot
12%
of all sites
8
sites block it
All 8 government websites that block GPTBot
| # | Site | Policy | Crawlers blocked |
|---|---|---|---|
| 275 | unesco.org | Blocks training only | 8 of 29 |
| 923 | meb.gov.tr | Blocks some AI search | 10 of 29 |
| 1216 | congress.gov | Blocks some AI search | 19 of 29 |
| 1509 | wa.gov | Blocks all AI search | 29 of 29 |
| 5924 | nv.gov | Blocks all AI search | 29 of 29 |
| 7487 | educationboardresults.gov.bd | Blocks all AI search | 29 of 29 |
| 8356 | ahrq.gov | Blocks some AI search | 3 of 29 |
| 8915 | vermont.gov | Blocks some AI search | 7 of 29 |
Popular government websites that allow GPTBot
Other AI crawlers and government websites
- ClaudeBot 3.8%
- PerplexityBot 3.3%
- Google-Extended 2.4%
- CCBot 3.8%
- Bytespider 3.8%
- Applebot-Extended 2.9%
- OAI-SearchBot 1.9%
GPTBot in other segments
- social websites 32.7%
- .fr (France) websites 31.8%
- news websites 31.5%
- reference websites 28.9%
- .it (Italy) websites 27.3%
- .nl (Netherlands) websites 27.1%
- .au (Australia) websites 27%
- .pl (Poland) websites 26.4%
- .es (Spain) websites 25%
- .de (Germany) websites 23.6%
- .uk (United Kingdom) websites 20%
- .app websites 16.7%
- .ch (Switzerland) websites 16.7%
- .net websites 13.6%
FAQ
- Which government websites block GPTBot?
- unesco.org, meb.gov.tr, congress.gov, wa.gov, nv.gov, educationboardresults.gov.bd, ahrq.gov, vermont.gov.
- How many government websites block GPTBot?
- As of 2026-10-03, 8 of the 210 government websites whose robots.txt we could read (3.8%) block GPTBot, OpenAI's AI training crawler, on the home page, compared with 12% of all sites.
- How do I block or allow GPTBot on my site?
- Add "User-agent: GPTBot" followed by "Disallow: /" to https://<your-site>/robots.txt to block it, or "Allow: /" to allow it. Check the result with the free AI crawler checker at tools.yukai.uk/free/ai-crawler-checker.