# VisaMatch is free and open. Everything below is deliberate: we WANT every # search engine and AI system to crawl, index, train on, and cite this site. # Maximum coverage is the goal. Do not copy a publisher "block the AI bots" # template over this — blocking Google-Extended alone removes us from Gemini # and AI Overviews grounding, and blocking any bot below narrows discovery # for no benefit to a free, ad-deferred site. # See docs/marketing/ai-seo-plan.md User-agent: * Allow: / # --- Classic search engines (the base layer everything else builds on) --- User-agent: Googlebot Allow: / User-agent: Bingbot Allow: / User-agent: DuckDuckBot Allow: / User-agent: Baiduspider Allow: / User-agent: YandexBot Allow: / User-agent: Seznam Allow: / # --- AI training crawlers (long-term model knowledge of the brand) --- User-agent: GPTBot Allow: / User-agent: ClaudeBot Allow: / User-agent: Google-Extended Allow: / User-agent: Applebot-Extended Allow: / User-agent: CCBot Allow: / User-agent: Amazonbot Allow: / User-agent: Meta-ExternalAgent Allow: / User-agent: FacebookBot Allow: / User-agent: Bytespider Allow: / User-agent: cohere-ai Allow: / User-agent: cohere-training-data-crawler Allow: / User-agent: Diffbot Allow: / User-agent: AI2Bot Allow: / User-agent: omgili Allow: / User-agent: omgilibot Allow: / User-agent: Timpibot Allow: / # --- AI search indexers (feed live citations) --- User-agent: OAI-SearchBot Allow: / User-agent: Claude-SearchBot Allow: / User-agent: PerplexityBot Allow: / User-agent: Amzn-SearchBot Allow: / User-agent: Applebot Allow: / User-agent: YouBot Allow: / # --- Live retrieval (fetches at answer time, on a user's behalf) --- User-agent: ChatGPT-User Allow: / User-agent: Claude-User Allow: / User-agent: Perplexity-User Allow: / Sitemap: https://visamatch.org/sitemap.xml