Agent Web Index › openmediavault.org

Can AI assistants read openmediavault.org?

Measured on 2026-09-19 by asking the site 9 times — once as a browser, once as each of the 8 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead.

A85 / 100
7 of 8 crawlers can read this page.
More readable than 72% of the 47,315 sites measured so far. Structured data: BreadcrumbList, Organization, WebPage, WebSite.

What each crawler got back

CrawlerHTTPResultrobots.txt
ClaudeBot (Claude)200 can read it no-rule
GPTBot (ChatGPT)200 can read it no-rule
OAI-SearchBot (ChatGPT Search)200 can read it no-rule
PerplexityBot (Perplexity)200 can read it no-rule
Google-Extended (Gemini, AI Overviews)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out no-rule
Meta-ExternalAgent (Meta AI)200 can read it no-rule
Amazonbot (Alexa, Rufus)200 allowed by the server, blocked in robots.txt blocked
Bytespider (Doubao, Lark)200 can read it no-rule
Applebot (Siri, Apple Intelligence)200 can read it no-rule
Applebot-Extended (Apple Intelligence training)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out no-rule

What to change, in order

  1. Publish sitemap.xml and llms.txt
    sitemap.xml and llms.txt are missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.
  2. 1 crawler is named and refused in robots.txt — a deliberate choice
    Amazonbot is blocked by a rule that names it. Nothing to fix here: this page records what is true, not what should be. It is listed so the deliberate part of the block is not confused with the accidental part above.

The checks

Re-run this audit live → Full AI-visibility report for openmediavault.org →

Sites with a similar score

marvelapp.comA 86openclaw.aiA 85openresty.comA 85onshape.comA 85ordergroove.comA 85nurd.comA 85papadontpreach.comA 85musickordi.comA 85pysznosci.plA 85hobbs.comA 85thetradedesk.comA 85pagesjaunes.caB 84

Other woocommerce stores

drivestrike.comA+ 93logix.inA 90hosterion.comA 86antidot.netB 79zdopravy.czC 71littlevectordesigns.comF 29

Browse the whole index →

Embed this score

Put the badge on openmediavault.org — it links back here, and re-measures every time this index re-crawls.

openmediavault.org AI readability: A 85/100
<a href="https://shop.lumnika.com/ai-readiness/openmediavault.org"><img src="https://shop.lumnika.com/ai-readiness/openmediavault.org/badge.svg" alt="AI readability"></a>
Method. 9 live HTTP requests (one per crawler, one as a browser) plus robots.txt, llms.txt, sitemap.xml and security.txt, 12-second timeout each, from a single vantage point. Blocking AI crawlers is a legitimate choice, not a failure: this page records what is true, not what should be. Domain comes from a shop-platform seed (Shopify/WooCommerce/Magento/BigCommerce/Shopline fingerprint or JSON-LD price), not the Tranco traffic ranking. One page per domain — the homepage — is audited.

Part of the Agent Web Index, 47,315 domains measured, updated continuously.