CrawlIndex reads publicly served files. It fetches robots.txt, homepages and two well-known paths, exactly as any crawler would. It stores no personal data, sets no cookies, and runs no analytics or third-party tracking.
To leave the index, disallow CrawlIndexBot in your robots.txt. It is checked before anything else is requested and takes effect on the next crawl.
MethodologyGlossaryCoverageDatasetAdd a domainCheck a domainAboutAPIllms.txtSource
Research and data by Fidget Labs BV, Breda, Netherlands. Free to reuse under CC BY 4.0 with attribution.