Faro Research
Faro scans websites for AI agent readiness all day, every day. These are the studies that come out of that index: how many sites block ChatGPT and Claude, how many publish structured data an AI system can actually use, and how far apart the industries are.
Every study states its sample, its method and what it cannot show. Every table comes with a CSV. Take the numbers and use them.
Faro runs a 47-check agent readiness scan across 7 categories and returns a score out of 100. It's the same scan behind the free AI Readiness Scan anyone can run on their own site, and the results feed the public Faro Score Index. The research is cut from that index, not from a separate crawl built to prove a point.
Sites enter the index through public company lists and a daily discovery job. Each one is fetched as an ordinary HTTP client with no JavaScript execution, no more than once a day, reading robots.txt, the homepage and its headers, and a short list of well-known paths. One row per registrable domain, newest scan wins.
That method has a real cost, and it's stated on every study: this is not a random sample of the web, and no figure here should be read as a web-wide rate. It's a large, consistently measured sample of commercially visible websites, measured the same way every time so the same measurement can be taken again.
Every study ships the table behind it as a CSV on a stable URL, and the page carries Dataset markup pointing at that file. The point is that the table can be checked and reused as data rather than screenshotted.
The tables are licensed CC BY 4.0. Republish them, chart them, quote the figures in an article or a deck. A link back to the study page is the only condition, and you don't need to ask first. Each study also carries a ready-made citation line with the sample size and fieldwork window in it, so the number and its caveat travel together.
Working on a story and need a cut that isn't on the page, by industry, by country or by individual check? Email hello@byfaro.ai with the question you are trying to answer and we will run it.
One study a quarter. These are the questions in the queue, not published findings: agent readiness by US state, so the geography of the gap is visible; the share of the largest SaaS companies publishing llms.txt and agents.json; and Fortune 500 against Inc 5000, to test whether size predicts readiness or works against it.
If there is a question you want answered from this index, say so. It's cheaper for us to run a query than for you to build the sample.