Does chatgpt.com block AI crawlers?
Blocks AI training and AI answers
chatgpt.com blocks both AI training crawlers and AI answer crawlers in its robots.txt.
Checked from chatgpt.com/robots.txt. Robots rules change; the live file is the source of truth.
- AI crawlers blocked
- 4 of 12
- AI bots named in robots.txt
- 5
- llms.txt
- Not found
- Popularity
- Tranco top 100 (#83)
Can AI assistants cite chatgpt.com?
- ChatGPT search: its search crawlers may read the site
- Claude: its search crawlers may read the site
- Perplexity: its search crawler is blocked
chatgpt.com opts out of AI training by Google, Common Crawl, ByteDance. Training crawlers and answer crawlers are separate: blocking training doesn't by itself remove a site from AI answers.
AI crawler access, bot by bot
| Crawler | Used for | Status on chatgpt.com |
|---|---|---|
| GPTBotOpenAI | AI training | Allowed |
| OAI-SearchBotOpenAI | AI answers | Allowed |
| ChatGPT-UserOpenAI | AI answers | Allowed |
| ClaudeBotAnthropic | AI training | Allowed |
| Claude-SearchBotAnthropic | AI answers | Allowed |
| Claude-UserAnthropic | AI answers | Allowed |
| PerplexityBotPerplexity | AI answers | Blocked |
| Google-ExtendedGoogle | AI training | Blocked |
| Applebot-ExtendedApple | AI training | Allowed |
| CCBotCommon Crawl | AI training | Blocked |
| BytespiderByteDance | AI training | Blocked |
| Meta-ExternalAgentMeta | AI training | Allowed |
How chatgpt.com compares
- 19% of the 1,278 popular sites we checked block GPTBot; chatgpt.com does not.
- 18% of the 1,278 popular sites we checked block ClaudeBot; chatgpt.com does not.
- 14% of the 1,278 popular sites we checked block PerplexityBot; chatgpt.com is one of them.
- Among 123 sites of similar popularity, 37% block at least one AI crawler.
The robots.txt rules that apply to AI crawlers
User-agent: CCBot Disallow: / User-agent: Google-Extended Disallow: / User-agent: anthropic-ai Disallow: / User-agent: Bytespider Disallow: / User-agent: PerplexityBot Disallow: / User-agent: * Allow: /$ Allow: /?* Allow: /api/share/og/ Allow: /g/ Allow: /s/ Allow: /gg/v/ Allow: /m/ Allow: /share/ Allow: /maps$ Allow: /maps/ Allow: /*/maps$ Allow: /*/maps/ Allow: /canvas/shared/ Allow: /*/images Allow: /images Allow: /*/library Allow: /library Allow: /favicon.ico Allow: /assets/favicon Allow: /cdn/assets/favicon Allow: /cdn/assets/ Allow: /auth/ Allow: /gpts$ Allow: /codex Allow: /*/codex Allow: /search$ Allow: /backend-anon/ Allow: /public-api/ Allow: /sitemap.xml Allow: /marketing-sitemap.xml Allow: /images-sitemap.xml Allow: /writing-tools-sitemap.xml Allow: /football-sitemap.xml Allow: /acquisition-landing-pages-sitemap.xml Allow: /100chats Allow: /api/public_content/ Allow: /backend-api/public_content/ Allow: /?ref=dotcom Allow: /overview Allow: /*/overview
Excerpt: only groups for all user-agents (*) or named AI crawlers, up to 40 rules per group.
Sites with a similar AI policy
Does chatgpt.com block GPTBot?
No. As of Oct 2, 2026, chatgpt.com's robots.txt doesn't block GPTBot.
Can ChatGPT search show results from chatgpt.com?
Its robots.txt doesn't block OAI-SearchBot or ChatGPT-User, so ChatGPT's search crawlers may read it. Whether pages are actually shown depends on ChatGPT.
Does chatgpt.com have an llms.txt file?
We didn't find a valid llms.txt at chatgpt.com/llms.txt when we checked.
Data: chatgpt.com's public robots.txt and llms.txt, fetched by AskableHQ. Popularity rank from the Tranco list (ID Y83YG). Methodology.