Measurement · By Asfand Yar Junejo · 2 min read

Ranking, retrieval, citation: why measure them separately?

Ranking, retrieval, mention, citation, recommendation and conversion are six different outcomes, and one does not guarantee the next. A page can rank without being retrieved by an AI assistant, be retrieved without being cited, and be cited without anyone clicking. Measuring them as one “visibility” number hides where the chain actually breaks.

What does each outcome mean?

Each one describes a different moment between a question being asked and a customer acting on the answer.

OutcomeWhat happened
RankingYour page appeared at a position in a list of results
RetrievalA system pulled your content into what it used to build an answer
MentionThe answer named your brand
CitationThe answer linked your page as a source
RecommendationThe answer suggested you as the choice
ConversionSomeone arrived and did what you wanted them to do

Which outcomes can actually be measured?

Rankings and conversions can be measured directly; mentions and citations can be observed by testing; retrieval mostly has to be inferred.

OutcomeHowCertainty
RankingSearch Console, rank trackersMeasured
RetrievalRarely exposed by platformsInferred
Mention / citationRepeated, logged prompt testsObserved
RecommendationDecision-stage prompt testsObserved
ReferralAnalytics referrer data, partlyPartly measured
ConversionGA4 or CRM eventsMeasured

How do you test AI answers properly?

Treat every prompt test as an experiment: fix the conditions, repeat it, log everything and compare like with like.

  1. Write prompt families, not single prompts: informational, comparison and decision-stage questions
  2. Record the conditions: platform, model if shown, date, region, signed in or not, search mode
  3. Separate cold and contextual tests: a fresh chat behaves differently from a long conversation
  4. Repeat each prompt several times, because answers vary run to run
  5. Log mentions, citations and recommendations as separate columns
  6. Re-run monthly and date every result, since platforms change without notice

What are the common measurement mistakes?

Treating one screenshot as a trend, and treating a correlation as a cause.

  • Reporting a single AI answer as proof of visibility
  • Adding mentions and citations together into one score
  • Crediting a change for a rise that also happened to competitors
  • Comparing results taken on different platforms or dates
  • Hiding tests that came back negative or inconclusive

The GEO Visibility Index in the Lab applies this routine to citation tracking.

Get the next note by email

Search Signals is the free newsletter version of this blog, with a two-page visibility checklist as a welcome gift.