Agent Web Index › xhonline.blog

Can AI assistants read xhonline.blog?

Measured on 2026-09-19 by asking the site 7 times — once as a browser, once as each of the 6 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead.

C68 / 100
5 of 6 crawlers can read this page.
More readable than 31% of the 47,315 sites measured so far. No JSON-LD.

What each crawler got back

CrawlerHTTPResultrobots.txt
ClaudeBot (Claude)200 can read it allowed-by-star
GPTBot (ChatGPT)200 can read it allowed-by-star
OAI-SearchBot (ChatGPT Search)200 can read it allowed-by-star
PerplexityBot (Perplexity)200 can read it allowed-by-star
Google-Extended (Gemini, AI Overviews)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out allowed-by-star
Meta-ExternalAgent (Meta AI)200 can read it allowed-by-star
Amazonbot (Alexa, Rufus)200 thin allowed-by-star

Cloudflare answers for this domain. Where a crawler above is refused while robots.txt allows it, the rule is applied by that edge, not written by the site — see how often each edge does this.

What to change, in order

  1. Let in the 1 crawler your own robots.txt already allows+1 crawler
    Amazonbot is turned away before reading the page (answered, but with far less text than a browser gets), while robots.txt permits it — so this block is not written in your site. Cloudflare answers for this domain. Cloudflare applies AI-crawler blocking at the edge, independently of robots.txt (Security → Bots, plus the managed robots.txt feature). https://developers.cloudflare.com/bots/concepts/bot/ Fixing it takes this domain from 5 to 6 of 6 crawlers.
  2. Declare the facts in JSON-LD
    There is no JSON-LD on the page, so every fact — who you are, what you sell, the price — has to be guessed out of the prose and the layout. A schema.org block is the difference between being quoted correctly and being paraphrased.
  3. Publish sitemap.xml and llms.txt
    sitemap.xml and llms.txt are missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.

The checks

Re-run this audit live → Full AI-visibility report for xhonline.blog →

Sites with a similar score

adashop.myshopify.comC 69cankaya.edu.trC 69one-link.myshopify.comC 69xfyun.cnC 68xiaohei.comC 68xanderbrown.myshopify.comC 68xn--90aivcdt6dxbc.xn--p1aiC 68wongcw.comC 68youwii.myshopify.comC 68veccie.myshopify.comC 68stateon.ruC 68europcar.comC 68

Browse the whole index →

Embed this score

Put the badge on xhonline.blog — it links back here, and re-measures every time this index re-crawls.

xhonline.blog AI readability: C 68/100
<a href="https://shop.lumnika.com/ai-readiness/xhonline.blog"><img src="https://shop.lumnika.com/ai-readiness/xhonline.blog/badge.svg" alt="AI readability"></a>
Method. 7 live HTTP requests (one per crawler, one as a browser) plus robots.txt, llms.txt, sitemap.xml and security.txt, 12-second timeout each, from a single vantage point. Blocking AI crawlers is a legitimate choice, not a failure: this page records what is true, not what should be. Domain comes from a shop-platform seed (Shopify/WooCommerce/Magento/BigCommerce/Shopline fingerprint or JSON-LD price), not the Tranco traffic ranking. One page per domain — the homepage — is audited.

Part of the Agent Web Index, 47,315 domains measured, updated continuously.