What do AI crawlers extract from your page?
An element-by-element inventory of what AI crawlers keep, discard, and only conditionally read — plus the parts you can exclude on purpose.
7 posts
An element-by-element inventory of what AI crawlers keep, discard, and only conditionally read — plus the parts you can exclude on purpose.
What happens to a page between fetch and answer, the eight signals answer-readiness turns on, and how to check every one of them by hand and at scale.
Three acronyms, one page. Here is what answer engines and generative search changed about optimization, and what they did not.
Every AI bot worth naming in your robots.txt, what each one actually controls, and which ones decide whether an assistant can find you at all.
Which AI bots exist, what each one actually controls, why robots.txt is only half the answer, and how to check whether assistants can reach your pages.
The highest-weighted answer-readiness signal is a passage that answers its own heading in one breath. What it is, how to write one, how to check it.
A walk through what actually gets extracted from a page when an AI system reads it, and which parts of your HTML never make the trip.