The Agent Web Index

How much of the web can AI assistants actually read?

Nobody knows, because nobody has asked the web itself. We do: every site in the Tranco ranking, fetched six times — once as a browser, then once as each of the crawlers behind ChatGPT, Claude, Perplexity and Gemini — and we publish exactly what came back. Every number on this page is counted from those requests. Nothing is modelled, sampled or estimated.

Selling online? See the Agent Commerce Index, the same audit filtered to confirmed stores →
Loading the index…

Who gets let in, and who stops them

For each assistant's crawler, the split that only a live request can reveal: sites that serve it, sites that ban it in robots.txt, and — the number that exists nowhere else — sites whose robots.txt says yes while the server or CDN says no. That last column is almost always an accident: the owner allowed the crawler and a bot rule blocks it anyway.

served blocked in robots.txt (a deliberate choice) blocked by the server or CDN despite robots.txt

Grades across the measured web

The same weighted audit used on a single site, applied to every domain in the index.

Look up any domain

If it is already in the index you get its page and its history. If it is not, you can measure it live right now.

The series