User-agent: * Allow: / # Sitemap location Sitemap: https://caspi.io/sitemap.xml # --------------------------------------------------------------------------- # AI & LLM crawlers — explicitly allowed so Caspi can be discovered, indexed # and cited by AI answer engines (ChatGPT, Gemini, Claude, Perplexity, etc.). # See also: https://caspi.io/llms.txt # --------------------------------------------------------------------------- # OpenAI (ChatGPT training, search & live browsing) User-agent: GPTBot Allow: / User-agent: OAI-SearchBot Allow: / User-agent: ChatGPT-User Allow: / # Google (Gemini / Vertex AI) User-agent: Google-Extended Allow: / # Anthropic (Claude) User-agent: ClaudeBot Allow: / User-agent: Claude-Web Allow: / User-agent: Claude-SearchBot Allow: / User-agent: anthropic-ai Allow: / # Apple Intelligence User-agent: Applebot-Extended Allow: / # Perplexity User-agent: PerplexityBot Allow: / User-agent: Perplexity-User Allow: / # Common Crawl (feeds many open LLMs) User-agent: CCBot Allow: / # Other AI assistants User-agent: Amazonbot Allow: / User-agent: Meta-ExternalAgent Allow: / User-agent: cohere-ai Allow: / User-agent: DuckAssistBot Allow: / # Social preview crawlers (link unfurling) User-agent: facebookexternalhit Allow: / User-agent: Twitterbot Allow: / User-agent: LinkedInBot Allow: / User-agent: Slackbot Allow: /