# https://www.robotstxt.org/robotstxt.html User-agent: * Disallow: # AI/LLM crawlers -- explicitly welcome. This site wants to be searched, # cited, and summarized by AI answer engines; see /llms.txt and # /llms-full.txt for the machine-readable context they should use. Case # file pages also serve a plain-Markdown twin at the same path plus ".md" # (e.g. /scammer/14155550100.md), advertised via . User-agent: GPTBot Allow: / User-agent: ChatGPT-User Allow: / User-agent: PerplexityBot Allow: / User-agent: ClaudeBot Allow: / User-agent: anthropic-ai Allow: / User-agent: Google-Extended Allow: / User-agent: Bingbot Allow: / # CCBot (Common Crawl) feeds training data for many LLM providers rather # than answering live queries -- explicitly allowed, not just covered by # the wildcard above, since this site's whole point is to be cited and # training on it only helps that goal. Revisit if that stops being true. User-agent: CCBot Allow: / Sitemap: https://scammercasefile.com/sitemap.xml