Tranco rank 1,898
clevelandclinic.org
clevelandclinic.org scores 79 out of 100 for agent readiness. It allows every answer-surface AI crawler.
Score 79 out of 100, grade B
- Last measured
- 2026-08-09 18:01 UTC
- First seen
- 2026-08-09
- Rubric / probe
- v1.0.0 / v2.0.0
Served through Cloudflare. Compare against everything else on the same stack.
How this score is made up
Agent access45 / 45
- Answer-surface crawlers allowed. All 11 answer-surface crawlers are allowed.
- Secondary crawlers allowed. All 12 secondary crawlers are allowed.
- Serves crawlers the same content. A crawler and a browser receive comparable responses.
Machine-readable surface16 / 25
- robots.txt published. robots.txt is present and parseable.
- Sitemap declared in robots.txt. robots.txt points crawlers at a sitemap.
- llms.txt published. /llms.txt is served but is off-spec: no H2 sections.
- agents.md published. No /agents.md.
Content structure18 / 30
- Organization schema. No Organization JSON-LD. Agents cannot reliably attribute this site to an entity.
- WebSite schema. No WebSite JSON-LD.
- Additional structured data. Also declares: MedicalOrganization, WebPage.
- Readable without JavaScript. 5,841 characters of text in the server response. Most crawlers do not execute JavaScript.
- Single top-level heading. 1 h1 element found.
- Semantic landmarks. Uses nav, header, footer, article.
Crawler policy
Read from clevelandclinic.org/robots.txt. A crawler is listed as blocked when the rules deny it the site root. This operator names no AI crawler explicitly.
| Crawler | Operator | Tier | Named | Status |
|---|---|---|---|---|
| GPTBot | OpenAI | 1 | no | Allowed |
| OAI-SearchBot | OpenAI | 1 | no | Allowed |
| ChatGPT-User | OpenAI | 1 | no | Allowed |
| ClaudeBot | Anthropic | 1 | no | Allowed |
| Claude-User | Anthropic | 1 | no | Allowed |
| Claude-SearchBot | Anthropic | 1 | no | Allowed |
| PerplexityBot | Perplexity | 1 | no | Allowed |
| Perplexity-User | Perplexity | 1 | no | Allowed |
| Google-Extended | 1 | no | Allowed | |
| Applebot-Extended | Apple | 1 | no | Allowed |
| meta-externalagent | Meta | 1 | no | Allowed |
| CCBot | Common Crawl | 2 | no | Allowed |
| Amazonbot | Amazon | 2 | no | Allowed |
| Bytespider | ByteDance | 2 | no | Allowed |
| cohere-ai | Cohere | 2 | no | Allowed |
| MistralAI-User | Mistral | 2 | no | Allowed |
| YouBot | You.com | 2 | no | Allowed |
| DuckAssistBot | DuckDuckGo | 2 | no | Allowed |
| kagi-fetcher | Kagi | 2 | no | Allowed |
| Diffbot | Diffbot | 2 | no | Allowed |
| Google-NotebookLM | 2 | no | Allowed | |
| TavilyBot | Tavily | 2 | no | Allowed |
| FirecrawlAgent | Firecrawl | 2 | no | Allowed |
Show this score
Free to embed. Always reflects the latest measurement, and links back here so anyone can check the working.
<a href="https://crawlindex.org/site/clevelandclinic.org"><img src="https://crawlindex.org/badge/clevelandclinic.org.svg" alt="CrawlIndex agent readiness score for clevelandclinic.org" width="196" height="28"></a>Using these figures
Free to reuse in research, journalism or a product under CC BY 4.0, with credit to Fidget Labs BV. Quote the measurement date so the claim stays checkable as the index moves.
CrawlIndex by Fidget Labs BV. "clevelandclinic.org agent readiness." https://crawlindex.org (measured 2026-08-09). Licensed CC BY 4.0.