← Research

The Crypto & Fintech AI-Readiness Index: Access Isn't the Problem

We audited 146 crypto and fintech sites for AI crawler access. 95.5% of the ones we could read allow every AI crawler we checked, and only 2 block a crawler that feeds a major AI answer engine. Whatever is keeping crypto and fintech brands out of AI answers, it is almost never a locked front door.

That matters because “you’re blocking AI crawlers by accident” has become the default GEO scare line. For these two verticals, the data says otherwise — and it points to where the real work is.

This is the first cut of a benchmark we’ll refresh quarterly at this URL: the Access dimension of our Retrieval Readiness Score, measured across the industry.

The headline numbers

Of 146 sampled domains, 132 returned a readable robots.txt (or none at all, which permits everything). Across those 132:

ResultSitesShare
Allow every AI crawler we checked12695.5%
Block at least one AI crawler64.5%
Block a primary answer-engine crawler (OpenAI / Anthropic / Perplexity)21.5%
Block Google-Extended (AI Overviews grounding)53.8%

The near-total openness holds across both verticals — but fintech is even more open than crypto:

VerticalSites readBlock anythingBlock an answer-engine crawler
Crypto (exchanges, DeFi, wallets, infra, L1/L2, media)775 (6.5%)2 (2.6%)
Fintech (neobanks, payments, lending, investing, infra)551 (1.8%)0

Not one of the 55 fintech sites we could read blocks a crawler that feeds the AI answer engines. The single fintech “blocker” (an infra company) blocks only training crawlers.

Even the blockers mostly aren’t blocking answers

Here’s the part the scare line misses. Of the six sites that block anything, five are making a content-licensing choice — refusing AI training crawlers (GPTBot, ClaudeBot, CCBot, Google-Extended, Bytespider, Applebot-Extended, meta-externalagent) while leaving the search/answer crawlers free to cite them.

DomainSegmentWhat it blocksRead
upbit.comCrypto exchangeEverything (all 13)The only true full block — invisible to every AI engine
coindesk.comCrypto mediaTraining bundle + PerplexityBotShuts out Perplexity’s answer crawler; ai-train=no otherwise
opensea.ioCrypto / DeFiTraining bundle + Amazonbotai-train=no; ChatGPT/Perplexity/Claude still allowed
gmx.ioCrypto / DeFiTraining bundle + Amazonbotai-train=no; answer engines still allowed
lithic.comFintech infraTraining bundleai-train=no; answer engines still allowed
trustwallet.comCrypto walletBytespider onlyBlocks one scraper; everything else allowed

So the true “you removed yourself from AI answers” count is two: upbit (which blocks everything) and CoinDesk (which blocks PerplexityBot). Everyone else who blocks is deliberately refusing training while staying citable — a defensible, intentional stance, not an accident.

What this means

If ~99% of crypto and fintech sites can be crawled by the answer engines, then the reason so few are actually cited isn’t access. The gap is downstream. Getting through the door is necessary, not sufficient. Being retrieved and named in an answer also requires:

  • Extraction — pages that break into clean, self-contained passages a model can lift.
  • Answerability — content that states the answer directly, early, with quotable specifics.
  • Entity Resolution — a brand the model can confidently recognize as a thing, not a string.

Those are the other three dimensions of the Retrieval Readiness Score. Access is the one most sites already pass — and the one the industry keeps talking about. The advantage is in the three nobody’s measuring at scale. You can check your own on all four in one pass with the free Retrieval Readiness Score tool.

Two caveats worth stating plainly

Robots.txt is not the whole access story. 14 of the 146 sites (about 10%) couldn’t be read at all — their edge firewalls returned 403/429 to our request, or they served an HTML app shell instead of a robots.txt. A WAF that blocks a plain robots.txt fetch from a datacenter IP may also be blocking real AI crawlers regardless of what robots.txt says — a failure mode robots.txt alone can’t detect. We count those as “unknown,” not “open.”

This is a single snapshot of one dimension. It measures root-path (/) access to a fixed list of 13 named crawlers on one day. It does not measure the three downstream dimensions, and robots.txt changes over time — which is exactly why we’ll re-run it quarterly at this URL.

Methodology

  • Sample: 146 crypto and fintech domains selected by primary review across exchanges, DeFi, wallets, on-chain infra, L1/L2 chains, crypto media, neobanks, payments, lending, investing, and fintech infrastructure — not scraped from a “best of” listicle.
  • Method: one fetch of each https://domain/robots.txt with a browser user-agent, parsed with the same RFC-9309 logic that powers our free AI Crawler Audit. For each of 13 named AI crawlers we resolve the most specific matching group and test whether / is disallowed. A missing robots.txt (404) is treated as allow-all.
  • Crawlers checked: GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, Claude-User, PerplexityBot, Perplexity-User, CCBot, Google-Extended, Applebot-Extended, meta-externalagent, Bytespider, Amazonbot.
  • Date: July 29, 2026. Reachable: 132 of 146. Raw data: the full per-domain results are published here.

This extends our earlier crypto-only crawler benchmark, which found a comparable ~2% block rate on a smaller set. Two snapshots, one story: access is rarely the barrier.


Want your own four-dimension score? Run the free Retrieval Readiness Score — no login — or book an audit and we’ll benchmark your domain against your category.


Zion Labs researches how AI answer engines choose what to cite, and builds the tools and monitoring to help brands become the source they trust. Try the free tools or book a free audit.