# A-Z Dates — open to all crawlers. The //// share-card # stubs are noindex,follow (set per-page), so they're intentionally NOT disallowed # here — blocking them would stop crawlers from seeing the noindex. User-agent: * Allow: / # ── Search engines we explicitly want ── # Naver (Yeti) — ~half the site is Korea content; Naver is the dominant KR engine. User-agent: Yeti Allow: / # ── Cooperate with AI crawlers — we want our curated content in AI overviews ── # Training/index crawlers User-agent: GPTBot Allow: / User-agent: ClaudeBot Allow: / User-agent: Google-Extended Allow: / User-agent: Applebot-Extended Allow: / User-agent: Amazonbot Allow: / User-agent: CCBot Allow: / User-agent: cohere-ai Allow: / User-agent: Meta-ExternalAgent Allow: / # Query-time search/citation agents (these fetch pages to cite them in answers) User-agent: OAI-SearchBot Allow: / User-agent: ChatGPT-User Allow: / User-agent: PerplexityBot Allow: / User-agent: Perplexity-User Allow: / User-agent: Claude-SearchBot Allow: / Sitemap: https://az.dates.ebenworks.co/sitemap-index.xml Sitemap: https://az.dates.ebenworks.co/sitemap.xml Sitemap: https://az.dates.ebenworks.co/sitemap-spokes.xml Sitemap: https://az.dates.ebenworks.co/sitemap-images.xml # EBENWORKS-AI-CRAWLERS:BEGIN — generated, do not edit between the markers # Source: infra/docs/ai-crawlers.json · sync: infra/scripts/sync-ai-crawlers.py # # Answer and search engines are allowed by name below, including the AI ones: # being cited in an answer is distribution. Training crawlers are denied by # name, and that refusal is the licence position — see the terms of use for # this site. Ignoring this file does not create a permission. User-agent: Googlebot User-agent: Googlebot-Image User-agent: Bingbot User-agent: DuckDuckBot User-agent: Applebot User-agent: YandexBot User-agent: Yeti User-agent: SeznamBot User-agent: OAI-SearchBot User-agent: ChatGPT-User User-agent: Claude-User User-agent: Claude-SearchBot User-agent: PerplexityBot User-agent: Perplexity-User Allow: / User-agent: GPTBot User-agent: ClaudeBot User-agent: anthropic-ai User-agent: CCBot User-agent: Google-Extended User-agent: Applebot-Extended User-agent: meta-externalagent User-agent: FacebookBot User-agent: Bytespider User-agent: Amazonbot User-agent: AI2Bot User-agent: Diffbot User-agent: omgili User-agent: omgilibot User-agent: ImagesiftBot User-agent: PanguBot User-agent: Timpibot User-agent: Kangaroo Bot User-agent: cohere-ai User-agent: Scrapy Disallow: / # EBENWORKS-AI-CRAWLERS:END