Agent Web Index › erpnext.com

Can AI assistants read erpnext.com?

Measured on 2026-09-19 by asking the site 9 times — once as a browser, once as each of the 8 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead.

C62 / 100
0 of 8 crawlers can read this page.
More readable than 20% of the 47,625 sites measured so far. Structured data: __invalid__, BreadcrumbList, FAQPage.

What each crawler got back

CrawlerHTTPResultrobots.txt
ClaudeBot (Claude) no-answer allowed-by-star
GPTBot (ChatGPT) no-answer allowed-by-star
OAI-SearchBot (ChatGPT Search) no-answer allowed-by-star
PerplexityBot (Perplexity) no-answer allowed-by-star
Google-Extended (Gemini, AI Overviews)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out allowed-by-star
Meta-ExternalAgent (Meta AI) no-answer allowed-by-star
Amazonbot (Alexa, Rufus) no-answer allowed-by-star
Bytespider (Doubao, Lark) no-answer allowed-by-star
Applebot (Siri, Apple Intelligence) no-answer allowed-by-star
Applebot-Extended (Apple Intelligence training)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out allowed-by-star

Cloudflare answers for this domain. Where a crawler above is refused while robots.txt allows it, the rule is applied by that edge, not written by the site — see how often each edge does this.

What to change, in order

  1. Let in the 8 crawlers your own robots.txt already allows+8 crawlers
    ClaudeBot, GPTBot, OAI-SearchBot, PerplexityBot, Meta-ExternalAgent, Amazonbot, Bytespider, Applebot are turned away before reading the page (the connection never completed), while robots.txt permits them — so this block is not written in your site. Cloudflare answers for this domain. Cloudflare applies AI-crawler blocking at the edge, independently of robots.txt (Security → Bots, plus the managed robots.txt feature). https://developers.cloudflare.com/bots/concepts/bot/ Fixing it takes this domain from 0 to 8 of 8 crawlers.
  2. Publish llms.txt
    llms.txt is missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.
  3. Mark the content with <main>
    The content-reachability check scores 30/100: with no <main> or <article> landmark, or with most of the page outside it, an assistant reads the navigation and the footer at the same cost as the text it came for.

Get told if this changes

One email only when a measured crawler flips on erpnext.com, served to refused or back. No schedule, no newsletter; double opt-in, one-click stop.

The checks

Re-run this audit live → Full AI-visibility report for erpnext.com →

Sites with a similar score

ijg.orgC 64airbnb.aeC 63eonline.comC 62esxdos.orgC 62eevblog.comC 62experiment.comC 62dl8.meC 62freedns.siC 62baawincasino.comC 62linksky.comC 62onestore.co.krD 61jappy.comD 59

Other woocommerce stores

cyberfolks.plA+ 93eurovisionled.comA 90futurehost.plA 86arcticwolf.comB 79zentrixads.comC 71littlevectordesigns.comF 29

Browse the whole index →

Embed this score

Put the badge on erpnext.com — it links back here, and re-measures every time this index re-crawls.

erpnext.com AI readability: C 62/100
<a href="https://shop.lumnika.com/ai-readiness/erpnext.com"><img src="https://shop.lumnika.com/ai-readiness/erpnext.com/badge.svg" alt="AI readability"></a>
Method. 9 live HTTP requests (one per crawler, one as a browser) plus robots.txt, llms.txt, sitemap.xml and security.txt, 12-second timeout each, from a single vantage point. Blocking AI crawlers is a legitimate choice, not a failure: this page records what is true, not what should be. Domain comes from a shop-platform seed (Shopify/WooCommerce/Magento/BigCommerce/Shopline fingerprint or JSON-LD price), not the Tranco traffic ranking. One page per domain — the homepage — is audited.

Part of the Agent Web Index, 47,625 domains measured, updated continuously.