Can AI assistants read soccerladuma.co.za?
Measured on 2026-09-19 by asking the site 7 times — once as a browser, once as each of the
6 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing
what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead.
C68 / 100
0 of 6 crawlers can read this page.
More readable than 31% of the 47,249 sites measured so far.
Structured data: NewsMediaOrganization, WebSite.
What each crawler got back
| Crawler | HTTP | Result | robots.txt |
|---|
| ClaudeBot (Claude) | 403 |
challenged |
blocked |
| GPTBot (ChatGPT) | 403 |
challenged |
blocked |
| OAI-SearchBot (ChatGPT Search) | 403 |
challenged |
allowed-by-star |
| PerplexityBot (Perplexity) | 403 |
challenged |
allowed-by-star |
| Google-Extended (Gemini, AI Overviews) a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own | — |
not opted out |
allowed-by-star |
| Meta-ExternalAgent (Meta AI) | 403 |
challenged |
blocked |
| Amazonbot (Alexa, Rufus) | 202 |
thin |
allowed-by-star |
AWS CloudFront answers for this domain. Where a crawler above is
refused while robots.txt allows it, the rule is applied by that edge, not written by the site — see
how often each edge does this.
What to change, in order
- Let in the 3 crawlers your own robots.txt already allows+3 crawlers
OAI-SearchBot, PerplexityBot, Amazonbot are turned away before reading the page (served a bot challenge instead of the page; answered, but with far less text than a browser gets), while robots.txt permits them — so this block is not written in your site. AWS CloudFront answers for this domain. AWS WAF Bot Control has an AI-crawler category that CloudFront enforces ahead of your origin. https://docs.aws.amazon.com/waf/latest/developerguide/waf-bot-control.html Fixing it takes this domain from 0 to 3 of 6 crawlers.
- 3 crawlers are named and refused in robots.txt — a deliberate choice
ClaudeBot, GPTBot, Meta-ExternalAgent are blocked by a rule that names them. Nothing to fix here: this page records what is true, not what should be. It is listed so the deliberate part of the block is not confused with the accidental part above.
Get told if this changes
One email only when a measured crawler flips on soccerladuma.co.za, served to refused or back. No schedule, no newsletter; double opt-in, one-click stop.
The checks
- Answers AI agents like it answers people14%
- Readable without JavaScript100%
- robots.txt lets the crawlers in57%
- Facts in JSON-LD84%
- Content reachable, not buried82%
- Publishes a map of itself100%
- Plain structure100%
Re-run this audit live →
Full AI-visibility report for soccerladuma.co.za →
Sites with a similar score
Browse the whole index →
Embed this score
Put the badge on soccerladuma.co.za — it links back here, and re-measures every time this index re-crawls.
<a href="https://shop.lumnika.com/ai-readiness/soccerladuma.co.za"><img src="https://shop.lumnika.com/ai-readiness/soccerladuma.co.za/badge.svg" alt="AI readability"></a>