Agent Web Index › hificorp.co.za

Can AI assistants read hificorp.co.za?

Measured on 2026-09-19 by asking the site 9 times — once as a browser, once as each of the 8 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead.

C62 / 100
2 of 8 crawlers can read this page.
More readable than 21% of the 47,315 sites measured so far. Structured data: WebSite, Organization, HowTo.

What each crawler got back

CrawlerHTTPResultrobots.txt
ClaudeBot (Claude)403 challenged blocked
GPTBot (ChatGPT)402 error blocked
OAI-SearchBot (ChatGPT Search)200 can read it allowed
PerplexityBot (Perplexity)402 error allowed
Google-Extended (Gemini, AI Overviews)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
opted out in robots.txt blocked
Meta-ExternalAgent (Meta AI)402 error blocked
Amazonbot (Alexa, Rufus)402 error blocked
Bytespider (Doubao, Lark)402 error blocked
Applebot (Siri, Apple Intelligence)200 can read it allowed
Applebot-Extended (Apple Intelligence training)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
opted out in robots.txt blocked

Cloudflare answers for this domain. Where a crawler above is refused while robots.txt allows it, the rule is applied by that edge, not written by the site — see how often each edge does this.

What to change, in order

  1. Let in the 1 crawler your own robots.txt already allows+1 crawler
    PerplexityBot is turned away before reading the page (answered with an HTTP error), while robots.txt permits it — so this block is not written in your site. Cloudflare answers for this domain. Cloudflare applies AI-crawler blocking at the edge, independently of robots.txt (Security → Bots, plus the managed robots.txt feature). https://developers.cloudflare.com/bots/concepts/bot/ Fixing it takes this domain from 2 to 3 of 8 crawlers.
  2. Publish llms.txt
    llms.txt is missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.
  3. Fix the plain structure: one <h1>, a title, a description, alt text
    The structure check scores 56/100. These are the cheapest signals on the page and the first ones an assistant uses to decide what the site is.
  4. 5 crawlers are named and refused in robots.txt — a deliberate choice
    ClaudeBot, GPTBot, Meta-ExternalAgent, Amazonbot, Bytespider are blocked by a rule that names them. Nothing to fix here: this page records what is true, not what should be. It is listed so the deliberate part of the block is not confused with the accidental part above.

The checks

Re-run this audit live → Full AI-visibility report for hificorp.co.za →

Sites with a similar score

lephoceen.frC 64bgtee.comC 63helsinginuutiset.fiC 62hike-summit.comC 62greenvideo.ioC 62htcdev.comC 62gameguardian.netC 62jup9.comC 62darkreader.orgC 62oke.zoneC 62safebrands.comD 61mbta.comD 59

Other magento stores

gymbeam.roB 84trafag.comB 80nkd.comB 75kasanova.comC 70bexley.frC 65eram.frD 46

Browse the whole index →

Embed this score

Put the badge on hificorp.co.za — it links back here, and re-measures every time this index re-crawls.

hificorp.co.za AI readability: C 62/100
<a href="https://shop.lumnika.com/ai-readiness/hificorp.co.za"><img src="https://shop.lumnika.com/ai-readiness/hificorp.co.za/badge.svg" alt="AI readability"></a>
Method. 9 live HTTP requests (one per crawler, one as a browser) plus robots.txt, llms.txt, sitemap.xml and security.txt, 12-second timeout each, from a single vantage point. Blocking AI crawlers is a legitimate choice, not a failure: this page records what is true, not what should be. Domain comes from a shop-platform seed (Shopify/WooCommerce/Magento/BigCommerce/Shopline fingerprint or JSON-LD price), not the Tranco traffic ranking. One page per domain — the homepage — is audited.

Part of the Agent Web Index, 47,315 domains measured, updated continuously.