Sitemap: https://www.leapfactor.io/sitemap.xml # AI crawlers, named explicitly. Two jobs share this group: retrieval bots # (OAI-SearchBot, ChatGPT-User, Claude-User, Perplexity-User) decide whether we # get cited in AI answers today, training bots (GPTBot, ClaudeBot, CCBot, # meta-externalagent) decide whether a future model recalls the brand at all. # A crawler obeys exactly one group, so an agent with no group of its own falls # through to "User-agent: *". Naming them here keeps a Disallow added down there # later from silently cutting them off -- and means the rules below reach them, # which the old "Allow: /" groups did not: those let AI bots crawl the admin # login and the uncached /blog?search= pages. User-agent: GPTBot User-agent: OAI-SearchBot User-agent: ChatGPT-User User-agent: ClaudeBot User-agent: Claude-User User-agent: Claude-SearchBot User-agent: Google-Extended User-agent: PerplexityBot User-agent: Perplexity-User User-agent: Applebot-Extended User-agent: Amazonbot User-agent: CCBot User-agent: meta-externalagent Disallow: /administrator Allow: /blog?page=* Allow: /blog?searchCategory=* Disallow: /blog?search=* # ByteDance's crawler feeds a China-market assistant we sell nothing into, and # it is a documented bandwidth hog. Stays out. User-agent: Bytespider Disallow: / User-agent: * Disallow: /administrator Allow: /blog?page=* Allow: /blog?searchCategory=* Disallow: /blog?search=*