Skip to content
AskableHQ
Launch

Does a1.net block AI crawlers?

Open to AI crawlers

a1.net doesn't block any of the major AI crawlers we track in its robots.txt.

Checked from a1.net/robots.txt. Robots rules change; the live file is the source of truth.

AI crawlers blocked
0 of 12
AI bots named in robots.txt
7
llms.txt
Not found
Popularity
Tranco top 5,000 (#3,023)

Can AI assistants cite a1.net?

  • ChatGPT search: its search crawlers may read the site
  • Claude: its search crawlers may read the site
  • Perplexity: its search crawlers may read the site

a1.net doesn't opt out of AI model training through robots.txt. Training crawlers and answer crawlers are separate: blocking training doesn't by itself remove a site from AI answers.

AI crawler access, bot by bot

CrawlerUsed forStatus on a1.net
GPTBotOpenAIAI trainingAllowed
OAI-SearchBotOpenAIAI answersAllowed
ChatGPT-UserOpenAIAI answersAllowed
ClaudeBotAnthropicAI trainingSome paths restricted
Claude-SearchBotAnthropicAI answersAllowed
Claude-UserAnthropicAI answersSome paths restricted
PerplexityBotPerplexityAI answersAllowed
Google-ExtendedGoogleAI trainingSome paths restricted
Applebot-ExtendedAppleAI trainingSome paths restricted
CCBotCommon CrawlAI trainingAllowed
BytespiderByteDanceAI trainingAllowed
Meta-ExternalAgentMetaAI trainingSome paths restricted

How a1.net compares

  • 19% of the 1,278 popular sites we checked block GPTBot; a1.net does not.
  • 18% of the 1,278 popular sites we checked block ClaudeBot; a1.net does not.
  • 14% of the 1,278 popular sites we checked block PerplexityBot; a1.net does not.
  • Among 170 sites of similar popularity, 26% block at least one AI crawler.

The robots.txt rules that apply to AI crawlers

User-agent: *
Disallow: /c/
Disallow: /delegate/
Disallow: /documents/
Disallow: /export/
Disallow: /group/
Disallow: /html/
Disallow: /o/
Disallow: /osgi/
Disallow: /portal/
Disallow: /user/
Disallow: /widget/
Disallow: /web/guest/
Disallow: /family-cube-250
Disallow: /familycube-250
Disallow: prico-2k18
Disallow: /checkout
Disallow: /warenkorb
Disallow: /bestellung
Disallow: /order
Disallow: /cart
Disallow: /search
Disallow: /suche
Disallow: /*?q=
Disallow: /*?keywords=
Disallow: /*?search=
Disallow: /*?query=
Disallow: /*?p_p_id=
Disallow: /*?p_l_id=
Disallow: /*?redirect=
Disallow: /*?doAsGroupId=
Disallow: /*?controlPanelCategory=
Disallow: /*&p_p_lifecycle=
Disallow: /*?p_p_lifecycle=
Disallow: /*?p_p_state=
Disallow: /*?p_p_mode=
Disallow: /*?inheritRedirect=
Disallow: /*?struts_action=
Disallow: /*?_com_liferay
Disallow: /*?cur=
Disallow: /*?delta=

User-agent: Claude-SearchBot
Allow: /

User-agent: GPTBot
Allow: /

User-agent: ChatGPT-User
Allow: /

User-agent: OAI-SearchBot
Allow: /

User-agent: CCBot
Allow: /

User-agent: PerplexityBot
Allow: /

User-agent: Bytespider
Allow: /

Excerpt: only groups for all user-agents (*) or named AI crawlers, up to 40 rules per group.

Sites with a similar AI policy

All sites that are “open to ai crawlers” →

Does a1.net block GPTBot?

No. As of Oct 2, 2026, a1.net's robots.txt doesn't block GPTBot.

Can ChatGPT search show results from a1.net?

Its robots.txt doesn't block OAI-SearchBot or ChatGPT-User, so ChatGPT's search crawlers may read it. Whether pages are actually shown depends on ChatGPT.

Does a1.net have an llms.txt file?

We didn't find a valid llms.txt at a1.net/llms.txt when we checked.

Data: a1.net's public robots.txt and llms.txt, fetched by AskableHQ. Popularity rank from the Tranco list (ID Y83YG). Methodology.