crawlindex

Search the index. A domain that is not here can still be measured live on the check page.

Tranco rank 3,692

abola.pt

abola.pt scores 77 out of 100 for agent readiness. It allows every answer-surface AI crawler. That is ahead of 87% of measured sites.

Score 77 out of 100, grade B

Agent Friendly

Number of measured sites in each ten-point score band
Score bandSites
0 to 90
10 to 199
20 to 2921
30 to 3950
40 to 49309
50 to 59733
60 to 69790
70 to 79466
80 to 89247
90 to 9913
Ahead of 87% of fully measured sites
Last measured
2026-08-17 03:28 UTC
First seen
2026-08-17
Rubric / probe
v2.0.0 / v3.0.0
VantageWhere the request came from. Origins serve differently by geography and IP reputation, so an observation is only comparable with another taken from the same place. Everything here is measured from GitHub runners in the US and EU. More
gha-ubuntu

Policy postureWhether anyone actually decided. Deliberate means robots.txt names AI crawlers by token. Inherited means it names none, so whatever AI policy exists is a side effect of generic rules. Blanket means one rule for everyone. Absent means no robots.txt at all. More

Inherited

robots.txt exists and names no AI crawler at all. Whatever AI policy this site has is a side effect of generic rules it inherited, most often from its platform or CDN default.

Access archetypeThe shape of the policy rather than its size. Open, no training, assistant only, selective, walled, metered, or undeclared. More

Open

Every answer-surface crawler is allowed. An agent asked about this site can read it.

Served through Cloudflare. Compare against everything else on the same stack.

How this score is made up

Points earned in each score band
BandEarnedAvailableNominal maximum
Agent access454545
Machine-readable surface72525
Content structure253030

Agent access45 / 45

  • Answer-surface crawlers allowed. All 11 answer-surface crawlers are allowed.
  • Secondary crawlers allowed. All 12 secondary crawlers are allowed.
  • Serves crawlers the same content. A crawler and a browser receive comparable responses.

Machine-readable surface7 / 25

  • robots.txt published. robots.txt is present and parseable.
  • Sitemap declared in robots.txt. robots.txt points crawlers at a sitemap.
  • llms.txt published. No /llms.txt.
  • agents.md published. No /agents.md.
  • Licence terms declared. No RSL License directive and no licence link relation. Reuse terms are undeclared.
  • Granular usage preferences declared. No Content-Signal directive. Policy is expressed only as allow or deny.
  • Agent card published. No agent card at /.well-known/agent-card.json.

Content structure25 / 30

  • Organization schema. Organization JSON-LD lets an agent resolve who publishes this site.
  • WebSite schema. WebSite JSON-LD present.
  • Additional structured data. Also declares: Person, PostalAddress, ImageObject, WebPage, MobileApplication.
  • Readable without JavaScript. 12,568 characters of text in the server response. Most crawlers do not execute JavaScript.
  • Single top-level heading. 0 h1 elements found.
  • Semantic landmarks. Uses main, nav, footer, article, aside.
  • Dateline declared. Declares a modification date in machine-readable form.
  • Authorship declared. No declared author. An agent quoting this page has nobody to credit.

How this compares to its peers

A score out of 100 is not information on its own. Against the sites running the same platform and sitting behind the same edge network, it is: this site is at or ahead of every cohort it belongs to, which usually means the ceiling is the platform rather than the site. Cohorts under 25 measured sites are never published, so every comparison here is against a real group.

This site's score against the median of each group it belongs to
GroupMedian scoreSites in groupDifference
This site7710
Sites behind Cloudflare64109513.0
The whole index61263816.0

What would move this score

The observable checks that did not earn full marks, heaviest first. Only these; a check we could not observe is not on the list, because it is not a fact about the site.

  1. +9llms.txt published. No /llms.txt.
  2. +4agents.md published. No /agents.md.
  3. +3Single top-level heading. 0 h1 elements found.
  4. +2Licence terms declared. No RSL License directive and no licence link relation. Reuse terms are undeclared.
  5. +2Granular usage preferences declared. No Content-Signal directive. Policy is expressed only as allow or deny.
  6. +2Authorship declared. No declared author. An agent quoting this page has nobody to credit.
  7. +1Agent card published. No agent card at /.well-known/agent-card.json.

Extraction profile

How cheap this page is for a retrieval pipeline to chunk. Recorded and published, deliberately not scored: it describes a shape rather than a pass or a fail, and compressing it into the grade would destroy the useful part.

Server text
12,568 chars
Text density
1%
Subheadings
2
Lists
4
Tables
0
Feed
yes

Crawler policy

Read from abola.pt/robots.txt. A crawler is listed as blocked when the rules deny it the site root. This operator names no AI crawler explicitly.

AI crawler access policy for abola.pt
CrawlerOperatorTierNamedStatus
GPTBotOpenAI1noAllowed
OAI-SearchBotOpenAI1noAllowed
ChatGPT-UserOpenAI1noAllowed
ClaudeBotAnthropic1noAllowed
Claude-UserAnthropic1noAllowed
Claude-SearchBotAnthropic1noAllowed
PerplexityBotPerplexity1noAllowed
Perplexity-UserPerplexity1noAllowed
Google-ExtendedGoogle1noAllowed
Applebot-ExtendedApple1noAllowed
meta-externalagentMeta1noAllowed
CCBotCommon Crawl2noAllowed
AmazonbotAmazon2noAllowed
BytespiderByteDance2noAllowed
cohere-aiCohere2noAllowed
MistralAI-UserMistral2noAllowed
YouBotYou.com2noAllowed
DuckAssistBotDuckDuckGo2noAllowed
kagi-fetcherKagi2noAllowed
DiffbotDiffbot2noAllowed
Google-NotebookLMGoogle2noAllowed
TavilyBotTavily2noAllowed
FirecrawlAgentFirecrawl2noAllowed

abola.pt has earned the Agent Friendly mark

Free to embed, no account and no fee. It regenerates from the nightly crawl, so it stays true, and it links back to this page so anyone can check the working in one click. What the mark means.

Shape

Circular. For a footer or an about page.

Theme

Follows the reader’s own light or dark setting. Note that this tracks the reader, not your page, so a dark-mode visitor sees the dark mark on a light site.

CrawlIndex agent readiness score for abola.ptOn a light page
On a dark page

The mark links back to this site's page for your domain, so anyone can check the claim in one click. It regenerates nightly, which means it stays true and it can change.

This measurement as JSON

crawl 2026-08-17 04:00 UTCprobe 3.0.0rubric 2.0.0registry 1.0.0vantage gha-ubuntu3,651 of 5,006 reachable

Using these figures

Free to reuse in research, journalism or a product under CC BY 4.0, with credit to Fidget Labs BV. Quote the measurement date so the claim stays checkable as the index moves.

CrawlIndex by Fidget Labs BV. "abola.pt agent readiness." https://crawlindex.org (measured 2026-08-17). Licensed CC BY 4.0.