Agent Web Index › thejns.org

Can AI assistants read thejns.org?

Measured on 2026-09-19 by asking the site 7 times — once as a browser, once as each of the 6 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead.

D58 / 100
1 of 6 crawlers can read this page.
More readable than 18% of the 47,821 sites measured so far. No JSON-LD.

What each crawler got back

CrawlerHTTPResultrobots.txt
ClaudeBot (Claude)403 challenged allowed
GPTBot (ChatGPT)403 challenged allowed
OAI-SearchBot (ChatGPT Search)403 challenged allowed-by-star
PerplexityBot (Perplexity)403 challenged allowed
Google-Extended (Gemini, AI Overviews)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out allowed
Meta-ExternalAgent (Meta AI)403 challenged allowed
Amazonbot (Alexa, Rufus)200 can read it allowed-by-star

AWS CloudFront answers for this domain. Where a crawler above is refused while robots.txt allows it, the rule is applied by that edge, not written by the site — see how often each edge does this.

What to change, in order

  1. Let in the 5 crawlers your own robots.txt already allows+5 crawlers
    ClaudeBot, GPTBot, OAI-SearchBot, PerplexityBot, Meta-ExternalAgent are turned away before reading the page (served a bot challenge instead of the page), while robots.txt permits them — so this block is not written in your site. AWS CloudFront answers for this domain. AWS WAF Bot Control has an AI-crawler category that CloudFront enforces ahead of your origin. https://docs.aws.amazon.com/waf/latest/developerguide/waf-bot-control.html Fixing it takes this domain from 1 to 6 of 6 crawlers.
  2. Declare the facts in JSON-LD
    There is no JSON-LD on the page, so every fact — who you are, what you sell, the price — has to be guessed out of the prose and the layout. A schema.org block is the difference between being quoted correctly and being paraphrased.
  3. Publish llms.txt
    llms.txt is missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.

Get told if this changes

One email only when a measured crawler flips on thejns.org, served to refused or back. No schedule, no newsletter; double opt-in, one-click stop.

The checks

Re-run this audit live → Full AI-visibility report for thejns.org →

Sites with a similar score

thelateblog.comD 6153kf.comD 59premierinn.comD 59thefp.comD 58therapyportal.comD 58tdesktop.comD 58tldp.orgD 58spantrix.comD 58univ-brest.frD 58oauth.netD 58zagreb.myshopify.comD 57shlog.myshopify.comD 57

Browse the whole index →

Embed this score

Put the badge on thejns.org — it links back here, and re-measures every time this index re-crawls.

thejns.org AI readability: D 58/100
<a href="https://shop.lumnika.com/ai-readiness/thejns.org"><img src="https://shop.lumnika.com/ai-readiness/thejns.org/badge.svg" alt="AI readability"></a>
Method. 7 live HTTP requests (one per crawler, one as a browser) plus robots.txt, llms.txt, sitemap.xml and security.txt, 12-second timeout each, from a single vantage point. Blocking AI crawlers is a legitimate choice, not a failure: this page records what is true, not what should be. Domain comes from a shop-platform seed (Shopify/WooCommerce/Magento/BigCommerce/Shopline fingerprint or JSON-LD price), not the Tranco traffic ranking. One page per domain — the homepage — is audited.

Part of the Agent Web Index, 47,821 domains measured, updated continuously.