User-agent: * Allow: / # The imprint, the privacy policy and the terms carry # and are deliberately left out of the # sitemap. # The wildcard above already allows everything below. The AI crawlers are named # anyway: "is this site open to answer engines" is a question asked about # robots.txt far more often than it is answered from one, and an explicit Allow # is the only way to say yes out loud. Both kinds are listed, the retrieval # crawlers that fetch a page to answer a question right now and the training # ones, because the blog exists to be quoted. User-agent: GPTBot Allow: / User-agent: OAI-SearchBot Allow: / User-agent: ChatGPT-User Allow: / User-agent: ClaudeBot Allow: / User-agent: Claude-User Allow: / User-agent: Claude-SearchBot Allow: / User-agent: anthropic-ai Allow: / User-agent: PerplexityBot Allow: / User-agent: Perplexity-User Allow: / User-agent: Google-Extended Allow: / User-agent: Applebot Allow: / User-agent: Applebot-Extended Allow: / User-agent: Bingbot Allow: / User-agent: DuckAssistBot Allow: / User-agent: meta-externalagent Allow: / User-agent: cohere-ai Allow: / Sitemap: https://clixad.io/sitemap.xml