Can AI assistants read sae.edu?
Measured on 2026-09-19 by asking the site 7 times — once as a browser, once as each of the
6 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing
what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead.
F43 / 100
1 of 6 crawlers can read this page.
More readable than 5% of the 47,315 sites measured so far.
The text is drawn by JavaScript, which most crawlers never run.
Structured data: WebSite.
What each crawler got back
| Crawler | HTTP | Result | robots.txt |
|---|
| ClaudeBot (Claude) | 403 |
challenged |
allowed-by-star |
| GPTBot (ChatGPT) | 403 |
challenged |
allowed-by-star |
| OAI-SearchBot (ChatGPT Search) | 200 |
can read it |
allowed-by-star |
| PerplexityBot (Perplexity) | 403 |
challenged |
allowed-by-star |
| Google-Extended (Gemini, AI Overviews) a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own | — |
not opted out |
allowed-by-star |
| Meta-ExternalAgent (Meta AI) | 403 |
challenged |
allowed-by-star |
| Amazonbot (Alexa, Rufus) | 403 |
challenged |
allowed-by-star |
What to change, in order
- Let in the 5 crawlers your own robots.txt already allows+5 crawlers
ClaudeBot, GPTBot, PerplexityBot, Meta-ExternalAgent, Amazonbot are turned away before reading the page (served a bot challenge instead of the page), while robots.txt permits them — so this block is not written in your site. No CDN signature was found in the response headers, so the refusal comes from the origin server itself or from a WAF this index does not recognise. Fixing it takes this domain from 1 to 6 of 6 crawlers.
- Put the text in the HTML, not only in JavaScript
The HTML that arrives is nearly empty and the content is painted by JavaScript. None of these crawlers run it, so even the ones that are served read a blank page. Server-rendering or prerendering the main content is what changes their result.
- Publish sitemap.xml and llms.txt
sitemap.xml and llms.txt are missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.
Get told if this changes
One email only when a measured crawler flips on sae.edu, served to refused or back. No schedule, no newsletter; double opt-in, one-click stop.
The checks
- Answers AI agents like it answers people29%
- Readable without JavaScript0%
- robots.txt lets the crawlers in100%
- Facts in JSON-LD67%
- Content reachable, not buried60%
- Publishes a map of itself0%
- Plain structure90%
Re-run this audit live →
Full AI-visibility report for sae.edu →
Sites with a similar score
Browse the whole index →
Embed this score
Put the badge on sae.edu — it links back here, and re-measures every time this index re-crawls.
<a href="https://shop.lumnika.com/ai-readiness/sae.edu"><img src="https://shop.lumnika.com/ai-readiness/sae.edu/badge.svg" alt="AI readability"></a>