Why the old metrics stopped working
Rank tracking assumes a stable, ordered list of results for a fixed query. AI answers have none of those properties: they are generated per-conversation, vary between users, and change with the phrasing of the prompt. A single number cannot describe your position because there is no position.
This is genuinely uncomfortable for teams used to a dashboard that goes up. The honest response is not to invent a fake ranking metric, but to measure the things that actually vary — whether you appear, how you are described, and whether you are recommended.
Track appearance rate, not position
Define a set of buying-intent prompts, run them regularly across the engines that matter, and record how often your brand appears at all. Appearance rate over a fixed prompt set is a stable, comparable number, and it moves when your work moves.
Keep the prompt set fixed for at least a quarter. Changing prompts and celebrating an improvement is the most common way teams fool themselves in this discipline.
Track how you are described
Appearing is not the same as appearing well. Capture the sentences the model uses about you and check them for accuracy, positioning, and completeness. A model that calls you a web design shop when you sell AI integration is a positioning failure that no amount of extra visibility fixes.
Being mentioned is table stakes. Being described the way you’d describe yourself is the actual goal.
Track recommendation, separately
Being listed among five options and being the recommended option are different outcomes with different causes. Score them separately. Recommendation tends to move with corroboration and specificity; mere inclusion tends to move with structure and markup.
Expect noise, and design for it
The same prompt asked twice can produce different answers, so a single check tells you almost nothing. Run each prompt several times, aggregate, and only treat a change as real if it survives repetition across a period. Week-to-week wobble is the medium, not a signal.
Set the review cadence to monthly. Anything faster and you will spend your time explaining variance to stakeholders instead of doing work that changes it.
Close the loop with the site
Every gap this measurement finds should convert into a specific page change: a question you never answered, a fact you left vague, a claim nothing corroborates. Measurement that doesn’t generate a backlog is theatre, and this discipline has enough of that already.