What AI search visibility actually measures
AI search visibility is whether your pages reach the answers that AI assistants write. Not whether you rank. Not whether you get traffic. Whether the words on your site end up inside the paragraph somebody reads when they ask an assistant about your category.
It is worth separating from two things it gets confused with. It is not the same as ranking, because an assistant reads a handful of results and writes from them, so being eleventh is functionally identical to being nowhere. And it is not the same as traffic, because most people who read an AI answer do not click a citation. They take the name and act on it later, which arrives in your analytics as a direct visit or a branded search with nothing connecting it back.
So visibility here is a thing you have to go and measure deliberately. Nothing arrives in a dashboard to tell you about it, and the absence of a signal is not evidence of absence.
There is a further wrinkle that catches people out. The same question asked of two different assistants can produce entirely different sets of brands, because they retrieve from different places and weigh what they find differently. So there is no single answer to whether you are visible in AI search. There are several answers, one per engine, and they can disagree sharply. A business can be a fixture in one and absent from another for months without anything being wrong on their end beyond which pool of results each engine happens to read.
You are trying to be in the pages an assistant reads, and quotable once it reads them. Those are two separate jobs with two different timescales.
What actually happens when someone asks
The mechanism matters because it tells you which of your problems is real. Four steps.
The assistant decides whether to search. Timeless questions get answered from what the model already absorbed. Current, commercial or specific questions trigger a live search. Almost anything involving buying triggers one, which is good news, because it means the answer comes from pages rather than from whatever the model happened to learn about your industry two years ago.
It writes its own search query. Rarely the words the person typed. Somebody asking what to get their sister who has just started running produces searches like best running gifts and beginner runner gear. The page that gets pulled in answers the machine's query, not the human's question.
It reads the top results. The first handful. Not page two. This is where ordinary search authority does its work, because the same ranking system that has always decided what is at the top still decides it.
It writes an answer and names sources. Pages containing a clear, self-contained, quotable claim get quoted. Pages containing marketing prose get skimmed past in favour of ones that do not.
Two separate failure points, then. Failing to be retrieved is an authority problem. Being retrieved and not quoted is a writing problem. They need different fixes, and most advice on this subject only addresses the second, which is why so much of it produces nothing for sites that have the first.
The authority half, which is slow
If your pages do not surface in the searches an assistant runs, nothing else on this page matters. This is the uncomfortable half because it does not respond to effort in the short term.
What decides whether you surface is the same thing that has always decided it: whether other sites reference you, how long your domain has existed, and whether you have genuine depth on the subject rather than one page about it. Topical authority is the useful concept here, and it is built by covering a subject thoroughly enough that a search engine treats you as a source on it rather than a page about it.
For a young site this is the binding constraint and pretending otherwise wastes months. A domain that is a few months old with almost nothing linking to it will not appear in competitive retrieved results regardless of how good its pages are. That is not a judgement on the writing. It is how the system works, and the honest response is to build depth while borrowing reach from places that already have it.
The practical version: pick a narrow subject you can genuinely own, build a real topic cluster around it rather than scattering single pages across a broad category, and accept that this compounds over quarters rather than weeks.
The mistake worth naming here is publishing width instead of depth. Twenty pages covering twenty loosely related subjects looks like a content programme and behaves like twenty orphans, because none of them is deep enough on anything for a search engine to treat the site as a source. Five pages that genuinely exhaust one narrow question, cross-referencing each other, do considerably better. This is unintuitive when the instinct is to cover more ground, and it is the single most common way a genuine content effort fails to move anything.
The quotable half, which is fast
This half responds immediately, costs nothing, and is where most sites have easy ground.
Lead with the claim. An assistant writing a three-sentence answer is looking for a sentence it can lift. Put the answer in the first line of a section and support it afterwards. A page that spends three paragraphs warming up gets skipped in favour of one that answers immediately.
Be specific enough to be checkable. Compare two versions of the same fact. We pride ourselves on fast shipping gives an assistant nothing. Orders placed before 2pm ship the same day, and we refund the delivery cost on anything that misses its window is a claim it can repeat and attribute. Specificity is what makes a sentence quotable.
Answer the question the page is named for. A page titled how to choose a filter for a 20 gallon tank should answer that in its first paragraph, not in section four after a history of aquarium filtration.
Structure so extraction is easy. Clear headings that describe what follows, short paragraphs, lists where a list is genuinely the shape of the answer, and FAQ markup on the questions. None of this is exotic. It is the same structure that makes a page readable by a person in a hurry.
A useful test before publishing anything: read your own page and try to find the one sentence an assistant would quote. If you cannot find it in under ten seconds, neither will the machine, and the page will lose to one where the claim is sitting in plain view. This test is unreasonably effective because it catches the most common problem in commercial writing, which is a page that circles a subject at length without ever committing to a statement somebody could disagree with.
The check that takes ten minutes and blocks everything
Before either half matters, the AI crawlers have to be able to read your site, and a surprising share of sites block them without anyone having decided to.
The ones worth allowing are GPTBot and OAI-SearchBot for OpenAI, ClaudeBot for Anthropic, PerplexityBot for Perplexity, and Google-Extended for Google's AI surfaces. Blocking any of them removes you from that engine's answers entirely.
Three things cause it in practice. A robots.txt file that disallows unfamiliar agents by default, which is a common inherited setting. A content delivery network with bot protection turned up high enough to challenge them. And a site that builds its content with JavaScript, since several of these crawlers read the raw HTML and find an empty page.
The reason this belongs first is that its outcome is binary and its fix is instant. Everything else on this page is a matter of degree. This one is a gate.
How long any of it takes
Different engines move on genuinely different timescales, and knowing which is which prevents a lot of premature despair.
Perplexity is the fastest. It retrieves live and leans heavily on current results, so a new page or a new mention can show up in its answers within days of being indexed. If you want early evidence that something is working, look here first.
ChatGPT is slower. Retrieval-driven change tends to appear over roughly ten to fourteen weeks. Anything that lives in what the model itself absorbed only changes when the model does, which is not on your schedule.
Google AI Overviews move with Google. They sit on top of ordinary rankings, so they shift when your rankings shift, which for a young site is measured in months. The AI Overviews surface is the one most tightly coupled to conventional search performance.
The reasonable expectation is a shift over one to two quarters of consistent work rather than a step change. That is worth saying plainly at the start, because the most common way this work gets abandoned is somebody checking after three weeks, seeing nothing, and concluding it does not work.
There is a second lag underneath the first that is easy to miss. Before an engine can retrieve your page, a search index has to have crawled and stored it, and for a young site that alone can take weeks. So the sequence is publish, wait to be indexed, wait to rank well enough to be in the handful an assistant reads, and only then appear in an answer. Three waits stacked on top of each other, none of which you control. It is worth tracking the earlier stages rather than only the final one, because indexation moving is the first honest sign that anything is happening at all.
How to know whether it is working
Because the traffic signal is unreliable, you have to measure the thing directly, and the discipline matters more than the tooling.
Ask buyer questions, not brand questions. Asking an assistant about your brand tells you what it knows about you. Asking what to buy tells you whether you exist at the moment of decision. Only the second is visibility.
Use the same set every time. Twenty questions, written in a buyer's words, frozen. Changing the wording restarts your history, so write them carefully once and keep the original somewhere separate.
Check at least two engines. They disagree constantly and the disagreement is informative. Present in Perplexity and absent from ChatGPT points at a different problem from absent everywhere.
Establish your noise floor before reading anything into a change. Run the whole set twice in one afternoon and see how much it moves on its own. Any change smaller than that gap carries no information, and most people are surprised how large it is.
Record who else gets named. This turns a pass or fail into a map of your category, and tells you who you are genuinely competing with for the mention, which is often not who you assumed. The checker guide covers running this properly.
A realistic order to work in
The order matters more than the list, because doing these out of sequence wastes the most time.
First, the crawler check. Ten minutes, binary, blocks everything downstream.
Second, write the twenty questions and take a baseline. Before changing anything. Without a baseline you cannot tell later whether anything moved, and the temptation to reconstruct one retrospectively produces a number nobody trusts.
Third, fix quotability on the pages you already have. This is the fast half. For every question where a competitor is named and you are not, look at whether you have a page that answers it and whether that page answers it in the first paragraph. Usually the page exists and buries the answer.
Fourth, and running in parallel from the start, build somewhere you can be found. Depth on a narrow subject on your own site, and genuine presence in the places that already get cited for your category. Both are slow, which is exactly why they start on day one rather than after the fast work is finished.
Fifth, re-run the identical question set at twelve weeks. Compare properly. Expect movement in Perplexity and very little in ChatGPT, because that is the shape of it. The optimization guide goes deeper on each lever.
Two jobs, two clocks. Quotability is free and fast and most sites have room there. Being retrievable is slow and is usually the real constraint. Work out which one is actually stopping you before spending a quarter on the other.