TidyToolsAI Crawler Index · data as of · updated daily

AI crawler access · congress.gov

Does congress.gov block GPTBot and other AI crawlers?

GPTBot: blockedAs of 2026-10-02, the robots.txt of congress.gov blocks 19 of 29 AI crawlers on the home page: GPTBot is blocked, ClaudeBot is blocked and PerplexityBot is blocked.

All 29 AI crawlers on congress.gov

Status for the home page (/) under congress.gov's robots.txt. Partially blocked = home page allowed, some paths disallowed.

CrawlerOperatorrobots.txt
Training crawlers
GPTBotOpenAIblocked
ClaudeBotAnthropicblocked
Google-ExtendedGoogleblocked
Applebot-ExtendedAppleallowed
Meta-ExternalAgentMetablocked
AmazonbotAmazonblocked
MistralAI-TrainingMistralallowed
CCBotCommon Crawlblocked
BytespiderByteDanceblocked
cohere-aiCohereblocked
AI search crawlers
OAI-SearchBotOpenAIblocked
Claude-SearchBotAnthropicblocked
PerplexityBotPerplexityblocked
Google-CloudVertexBotGoogleblocked
ApplebotAppleallowed
meta-webindexerMetablocked
Amzn-SearchBotAmazonallowed
DuckAssistBotDuckDuckGoblocked
MistralAI-IndexMistralallowed
User-triggered fetch crawlers
ChatGPT-UserOpenAIblocked
Claude-UserAnthropicblocked
Perplexity-UserPerplexityblocked
Google-GeminiNotebookGoogleallowed
Google-AgentGoogleallowed
meta-externalfetcherMetablocked
Amzn-UserAmazonallowed
MistralAI-UserMistralblocked
Ads review crawlers
OAI-AdsBotOpenAIallowed
meta-externaladsMetaallowed

Compared with other government sites in the index: 3.8% of them block GPTBot and 3.8% block ClaudeBot (n=210). Across all 8,366 readable sites: 12% block GPTBot. All government sites →

On .gov domains: 4.6% of 108 readable sites block GPTBot and 1.9% serve an llms.txt. All .gov sites →

History

No change recorded since daily tracking started on 2026-10-02. We re-check congress.gov every day; a change in its AI crawler rules shows up here and on the index page.

FAQ

Does congress.gov block GPTBot?
Yes. As of 2026-10-02, https://congress.gov/robots.txt disallows the home page for GPTBot, OpenAI's AI training crawler.
Does congress.gov block ClaudeBot?
Yes. As of 2026-10-02, https://congress.gov/robots.txt disallows the home page for ClaudeBot, Anthropic's AI training crawler.
Does congress.gov block PerplexityBot?
Yes. As of 2026-10-02, https://congress.gov/robots.txt disallows the home page for PerplexityBot, Perplexity's AI search crawler.
Can ChatGPT search (OAI-SearchBot) crawl congress.gov?
Yes. As of 2026-10-02, https://congress.gov/robots.txt disallows the home page for OAI-SearchBot, OpenAI's AI search crawler.
Does congress.gov block Google-Extended (Gemini training)?
Yes. As of 2026-10-02, https://congress.gov/robots.txt disallows the home page for Google-Extended, Google's AI training crawler.
Does congress.gov have an llms.txt file?
No. As of 2026-10-02, congress.gov does not serve a plain-text /llms.txt file.
How can I check congress.gov or my own site again?
Use the free AI crawler checker at tools.yukai.uk/free/ai-crawler-checker (no sign-up). To check many sites and get alerts when rules change, use the AI Crawler Access Checker on Apify.

Similar government sites

Other government sites block GPTBot · Which .gov sites block GPTBot

All 10,000 sites · All 29 AI crawlers · AI Crawler Index