Skip to content
AskableHQ
Launch

Does slate.com block AI crawlers?

Blocks AI training, allows AI answers

slate.com blocks at least one AI training crawler but still lets AI answer engines read its pages.

Checked from slate.com/robots.txt. Robots rules change; the live file is the source of truth.

AI crawlers blocked
7 of 12
AI bots named in robots.txt
7
llms.txt
Not found
Popularity
Tranco top 5,000 (#2,500)

Can AI assistants cite slate.com?

  • ChatGPT search: its search crawlers may read the site
  • Claude: its search crawlers may read the site
  • Perplexity: its search crawlers may read the site

slate.com opts out of AI training by OpenAI, Anthropic, Google, Apple, Common Crawl, ByteDance, Meta. Training crawlers and answer crawlers are separate: blocking training doesn't by itself remove a site from AI answers.

AI crawler access, bot by bot

CrawlerUsed forStatus on slate.com
GPTBotOpenAIAI trainingBlocked
OAI-SearchBotOpenAIAI answersSome paths restricted
ChatGPT-UserOpenAIAI answersSome paths restricted
ClaudeBotAnthropicAI trainingBlocked
Claude-SearchBotAnthropicAI answersSome paths restricted
Claude-UserAnthropicAI answersSome paths restricted
PerplexityBotPerplexityAI answersSome paths restricted
Google-ExtendedGoogleAI trainingBlocked
Applebot-ExtendedAppleAI trainingBlocked
CCBotCommon CrawlAI trainingBlocked
BytespiderByteDanceAI trainingBlocked
Meta-ExternalAgentMetaAI trainingBlocked

How slate.com compares

  • 19% of the 1,278 popular sites we checked block GPTBot; slate.com is one of them.
  • 18% of the 1,278 popular sites we checked block ClaudeBot; slate.com is one of them.
  • 14% of the 1,278 popular sites we checked block PerplexityBot; slate.com does not.
  • Among 199 sites of similar popularity, 19% block at least one AI crawler.

The robots.txt rules that apply to AI crawlers

User-agent: *
Disallow: /search
Disallow: /comments/
Disallow: /_

User-agent: 008
User-agent: AdsTxtCrawler
User-agent: AgentDataBot
User-agent: AI2Bot
User-agent: Ai2Bot-Dolma
User-agent: Amazonbot
User-agent: ApifyBot
User-agent: ApifyWebsiteContentCrawler
User-agent: Applebot-Extended
User-agent: BeansLLM-CorpusBot
User-agent: bedrockbot
User-agent: BixelBot
User-agent: Bravebot
User-agent: Brightbot
User-agent: Bytespider
User-agent: CCBot
User-agent: ChatGLM-Spider
User-agent: ClaudeBot
User-agent: CloudflareBrowserRenderingCrawler
User-agent: CloudVertexBot
User-agent: cohere-training-data-crawler
User-agent: CohereBot
User-agent: Cotoyogi
User-agent: CragCrawler
User-agent: Crawlspace
User-agent: Datenbank Crawler
User-agent: dcrawl
User-agent: DeepSeekBot
User-agent: DesearchBot
User-agent: Diffbot
User-agent: Doubaobot
User-agent: ERNIEBot
User-agent: ExaBot
User-agent: FacebookBot
User-agent: FirecrawlAgent
User-agent: Google-Extended
User-agent: GoogleOther
User-agent: GPTBot
User-agent: HelloworkJobPostingBot
User-agent: HenkBot
User-agent: HTTrack
User-agent: HTTrack 3.0
User-agent: Hunyuan
User-agent: ICC-Crawler
User-agent: imageSpider
User-agent: IndeedJobBot
User-agent: Kangaroo Bot
User-agent: Keenable-User
User-agent: KeenableBot
User-agent: KimiBot
User-agent: KimiCrawler
User-agent: KrawlerBot
User-agent: laion-huggingface-processor
User-agent: LCC
User-agent: meta-externalagent
User-agent: MetaInspector
User-agent: MetaJobBot
User-agent: micro-crawl
User-agent: MistralBot
User-agent: MoonshotBot
User-agent: MoonSpider
User-agent: Mozilla-Tabstack
User-agent: netEstate Imprint Crawler
User-agent: newspaper
User-agent: Nutch
User-agent: nyt_scraping
User-agent: Offline Explorer
User-agent: omgili
User-agent: OpenindexSpider
User-agent: PanguBot
User-agent: Potions
User-agent: Querit-SearchBot
User-agent: QueritBot
User-agent: QwenBot
User-agent: recraft-image-dataset
User-agent: ReflectionBot
User-agent: SalamandraVLM
User-agent: SBIntuitionsBot
User-agent: Scrapy
User-agent: ScryBot
User-agent: ServerHunterSpider
User-agent: ShapBot
User-agent: SofyaBot
User-agent: Spider
User-agent: StatsDroneBot
User-agent: TalarionFrontierCrawler
User-agent: TavilyBot
User-agent: Terra Cotta
User-agent: TerraCotta
User-agent: Thinkbot
User-agent: Timpibot
User-agent: VelenPublicWebCrawler
User-agent: web-image-collection-research
User-agent: webzio-extended
User-agent: YiBot
User-agent: YouBot
Disallow: /

Excerpt: only groups for all user-agents (*) or named AI crawlers, up to 40 rules per group.

Sites with a similar AI policy

All sites that are “blocks ai training, allows ai answers” →

Does slate.com block GPTBot?

Yes. As of Oct 2, 2026, slate.com's robots.txt disallows GPTBot, OpenAI's training crawler.

Can ChatGPT search show results from slate.com?

Its robots.txt doesn't block OAI-SearchBot or ChatGPT-User, so ChatGPT's search crawlers may read it. Whether pages are actually shown depends on ChatGPT.

Does slate.com have an llms.txt file?

We didn't find a valid llms.txt at slate.com/llms.txt when we checked.

Data: slate.com's public robots.txt and llms.txt, fetched by AskableHQ. Popularity rank from the Tranco list (ID Y83YG). Methodology.