Tranco rank 4,059
funinexch.com
funinexch.com scores 61 out of 100 for agent readiness. It allows every answer-surface AI crawler. That is ahead of 50% of measured sites.
Measured
| Score band | Sites |
|---|---|
| 0 to 9 | 0 |
| 10 to 19 | 9 |
| 20 to 29 | 21 |
| 30 to 39 | 50 |
| 40 to 49 | 309 |
| 50 to 59 | 733 |
| 60 to 69 | 790 |
| 70 to 79 | 466 |
| 80 to 89 | 247 |
| 90 to 99 | 13 |
- Last measured
- 2026-08-17 03:37 UTC
- First seen
- 2026-08-17
- Rubric / probe
- v2.0.0 / v3.0.0
Vantage
Where the request came from. Origins serve differently by geography and IP reputation, so an observation is only comparable with another taken from the same place. Everything here is measured from GitHub runners in the US and EU. More- gha-ubuntu
Policy posture
Whether anyone actually decided. Deliberate means robots.txt names AI crawlers by token. Inherited means it names none, so whatever AI policy exists is a side effect of generic rules. Blanket means one rule for everyone. Absent means no robots.txt at all. More
Policy posture
Whether anyone actually decided. Deliberate means robots.txt names AI crawlers by token. Inherited means it names none, so whatever AI policy exists is a side effect of generic rules. Blanket means one rule for everyone. Absent means no robots.txt at all. MoreInherited
robots.txt exists and names no AI crawler at all. Whatever AI policy this site has is a side effect of generic rules it inherited, most often from its platform or CDN default.
Access archetype
The shape of the policy rather than its size. Open, no training, assistant only, selective, walled, metered, or undeclared. More
Access archetype
The shape of the policy rather than its size. Open, no training, assistant only, selective, walled, metered, or undeclared. MoreOpen
Every answer-surface crawler is allowed. An agent asked about this site can read it.
Built on Angular. Compare against everything else on the same stack.
This site's stated policy is not the one being enforced
robots.txt permits GPTBot and the server refuses GPTBot anyway. The operator published one policy and a different one is being enforced, almost always by an edge rule switched on above them. More
This site's stated policy is not the one being enforced
robots.txt permits GPTBot and the server refuses GPTBot anyway. The operator published one policy and a different one is being enforced, almost always by an edge rule switched on above them. Morerobots.txt permits GPTBot, and a request identifying as GPTBot was refused with HTTP 403. Nothing in robots.txt asked for that, so it is almost certainly an edge rule applied above the operator rather than a decision they made. It is worth knowing about either way, because agents experience the enforcement and not the file.
How this score is made up
| Band | Earned | Available | Nominal maximum |
|---|---|---|---|
| Agent access | 38 | 45 | 45 |
| Machine-readable surface | 7 | 25 | 25 |
| Content structure | 16 | 30 | 30 |
Agent access38 / 45
- Answer-surface crawlers allowed. All 11 answer-surface crawlers are allowed.
- Secondary crawlers allowed. All 12 secondary crawlers are allowed.
- Serves crawlers the same content. Requesting as GPTBot returned HTTP 403 and 118 bytes, against 35,345 bytes for a browser.
Machine-readable surface7 / 25
- robots.txt published. robots.txt is present and parseable.
- Sitemap declared in robots.txt. robots.txt points crawlers at a sitemap.
- llms.txt published. No /llms.txt.
- agents.md published. No /agents.md.
- Licence terms declared. No RSL License directive and no licence link relation. Reuse terms are undeclared.
- Granular usage preferences declared. No Content-Signal directive. Policy is expressed only as allow or deny.
- Agent card published. No agent card at /.well-known/agent-card.json.
Content structure16 / 30
- Organization schema. No Organization JSON-LD. Agents cannot reliably attribute this site to an entity.
- WebSite schema. WebSite JSON-LD present.
- Additional structured data. Also declares: SportsOrganization.
- Readable without JavaScript. 13,604 characters of text in the server response. Most crawlers do not execute JavaScript.
- Single top-level heading. 1 h1 element found.
- Semantic landmarks. No semantic landmark elements found.
- Dateline declared. No machine-readable date. An agent cannot tell how current this page is.
- Authorship declared. No declared author. An agent quoting this page has nobody to credit.
How this compares to its peers
A score out of 100 is not information on its own. Against the sites running the same platform and sitting behind the same edge network, it is: this site is at or ahead of every cohort it belongs to, which usually means the ceiling is the platform rather than the site. Cohorts under 25 measured sites are never published, so every comparison here is against a real group.
| Group | Median score | Sites in group | Difference |
|---|---|---|---|
| This site | 61 | 1 | 0 |
| The whole index | 61 | 2638 | 0.0 |
What would move this score
The observable checks that did not earn full marks, heaviest first. Only these; a check we could not observe is not on the list, because it is not a fact about the site.
- +9llms.txt published. No /llms.txt.
- +7Serves crawlers the same content. Requesting as GPTBot returned HTTP 403 and 118 bytes, against 35,345 bytes for a browser.
- +7Organization schema. No Organization JSON-LD. Agents cannot reliably attribute this site to an entity.
- +4agents.md published. No /agents.md.
- +3Dateline declared. No machine-readable date. An agent cannot tell how current this page is.
- +2Licence terms declared. No RSL License directive and no licence link relation. Reuse terms are undeclared.
- +2Granular usage preferences declared. No Content-Signal directive. Policy is expressed only as allow or deny.
- +2Semantic landmarks. No semantic landmark elements found.
Extraction profile
How cheap this page is for a retrieval pipeline to chunk. Recorded and published, deliberately not scored: it describes a shape rather than a pass or a fail, and compressing it into the grade would destroy the useful part.
- Server text
- 13,604 chars
- Text density
- 38%
- Subheadings
- 48
- Lists
- 8
- Tables
- 0
- Feed
- no
Crawler policy
Read from funinexch.com/robots.txt. A crawler is listed as blocked when the rules deny it the site root. This operator names no AI crawler explicitly.
| Crawler | Operator | Tier | Named | Status |
|---|---|---|---|---|
| GPTBot | OpenAI | 1 | no | Allowed |
| OAI-SearchBot | OpenAI | 1 | no | Allowed |
| ChatGPT-User | OpenAI | 1 | no | Allowed |
| ClaudeBot | Anthropic | 1 | no | Allowed |
| Claude-User | Anthropic | 1 | no | Allowed |
| Claude-SearchBot | Anthropic | 1 | no | Allowed |
| PerplexityBot | Perplexity | 1 | no | Allowed |
| Perplexity-User | Perplexity | 1 | no | Allowed |
| Google-Extended | 1 | no | Allowed | |
| Applebot-Extended | Apple | 1 | no | Allowed |
| meta-externalagent | Meta | 1 | no | Allowed |
| CCBot | Common Crawl | 2 | no | Allowed |
| Amazonbot | Amazon | 2 | no | Allowed |
| Bytespider | ByteDance | 2 | no | Allowed |
| cohere-ai | Cohere | 2 | no | Allowed |
| MistralAI-User | Mistral | 2 | no | Allowed |
| YouBot | You.com | 2 | no | Allowed |
| DuckAssistBot | DuckDuckGo | 2 | no | Allowed |
| kagi-fetcher | Kagi | 2 | no | Allowed |
| Diffbot | Diffbot | 2 | no | Allowed |
| Google-NotebookLM | 2 | no | Allowed | |
| TavilyBot | Tavily | 2 | no | Allowed |
| FirecrawlAgent | Firecrawl | 2 | no | Allowed |
No mark at this score
The embeddable mark starts at grade B, because offering a graphic nobody would put on their own site is a pretence rather than a feature. The list above is the useful version: for most sites the points are concentrated in two or three changes. How the mark works.
The neutral mark above exists and is free to use if you want to show that the site is independently measured whatever the number says.
crawl 2026-08-17 04:00 UTCprobe 3.0.0rubric 2.0.0registry 1.0.0vantage gha-ubuntu3,651 of 5,006 reachable
Using these figures
Free to reuse in research, journalism or a product under CC BY 4.0, with credit to Fidget Labs BV. Quote the measurement date so the claim stays checkable as the index moves.
CrawlIndex by Fidget Labs BV. "funinexch.com agent readiness." https://crawlindex.org (measured 2026-08-17). Licensed CC BY 4.0.